Methods › General › Position Embeddings › Absolute Position Encodings › Papers, page 135
Absolute Position Encodings
Papers archive 2025-07-28
archive papers tagged: 13,942 · with a code link: 6,505 · where Syntology ran a sample: 2,224 (1,897 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,224 of 13,942 tagged: 1,897 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 135 of 140: papers 13,401 to 13,500 of 13,942, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Improving Generalization of Transformer for Speech Recognition with Parallel Schedule Sampling and Relative Positional Embedding 1 Nov 2019 · 0 repositories · arXiv:1911.00203
-
Improving Natural Language Understanding by Reverse Mapping Bytepair Encoding 1 Nov 2019 · 0 repositories
-
Inspecting Unification of Encoding and Matching with Transformer: A Case Study of Machine Reading Comprehension 1 Nov 2019 · 0 repositories
-
Long Warm-up and Self-Training: Training Strategies of NICT-2 NMT System at WAT-2019 1 Nov 2019 · 0 repositories
-
LTRC-MT Simple & Effective Hindi-English Neural Machine Translation Systems at WAT 2019 1 Nov 2019 · 0 repositories
-
Mixed Multi-Head Self-Attention for Neural Machine Translation 1 Nov 2019 · 0 repositories
-
On the Relation between Position Information and Sentence Length in Neural Machine Translation 1 Nov 2019 · 0 repositories
-
Our Neural Machine Translation Systems for WAT 2019 1 Nov 2019 · 0 repositories
-
Recurrent Positional Embedding for Neural Machine Translation 1 Nov 2019 · 0 repositories
-
Recycling a Pre-trained BERT Encoder for Neural Machine Translation 1 Nov 2019 · 0 repositories
-
Sarah's Participation in WAT 2019 1 Nov 2019 · 0 repositories
-
Self-Adaptive Scaling for Learnable Residual Structure 1 Nov 2019 · 0 repositories
-
Supervised neural machine translation based on data augmentation and improved training & inference process 1 Nov 2019 · 0 repositories
-
SYSTRAN @ WAT 2019: Russian-Japanese News Commentary task 1 Nov 2019 · 0 repositories
-
SYSTRAN @ WNGT 2019: DGT Task 1 Nov 2019 · 0 repositories
-
The Concordia NLG Surface Realizer at SRST 2019 1 Nov 2019 · 0 repositories
-
Transformer and seq2seq model for Paraphrase Generation 1 Nov 2019 · 0 repositories
-
Transformer-based Model for Single Documents Neural Summarization 1 Nov 2019 · 0 repositories
-
Transformer Dissection: An Unified Understanding for Transformer's Attention via the Lens of Kernel 1 Nov 2019 · 0 repositories
-
``Transforming'' Delete, Retrieve, Generate Approach for Controlled Text Style Transfer 1 Nov 2019 · 0 repositories
-
Attention Is All You Need for Chinese Word Segmentation 31 Oct 2019 · 1 repository · arXiv:1910.14537
-
Document-level Neural Machine Translation with Associated Memory Network 31 Oct 2019 · 0 repositories · arXiv:1910.14528
-
Image-Conditioned Graph Generation for Road Network Extraction 31 Oct 2019 · 3 repositories · arXiv:1910.14388
-
NAT: Neural Architecture Transformer for Accurate and Compact Architectures 31 Oct 2019 · 1 repository · arXiv:1910.14488
-
Neural Assistant: Joint Action Prediction, Response Generation, and Latent Knowledge Reasoning 31 Oct 2019 · 1 repository · arXiv:1910.14613
-
Parameter Sharing Decoder Pair for Auto Composing 31 Oct 2019 · 0 repositories · arXiv:1910.14270
-
Transfer Learning from Transformers to Fake News Challenge Stance Detection (FNC-1) Task 31 Oct 2019 · 0 repositories · arXiv:1910.14353
-
An Augmented Transformer Architecture for Natural Language Generation Tasks 30 Oct 2019 · 0 repositories · arXiv:1910.13634
-
Lightweight and Efficient End-to-End Speech Recognition Using Low-Rank Transformer 30 Oct 2019 · 0 repositories · arXiv:1910.13923
-
Transformer-based Cascaded Multimodal Speech Translation 29 Oct 2019 · 0 repositories · arXiv:1910.13215
-
An Empirical Study of Generation Order for Machine Translation 29 Oct 2019 · 0 repositories · arXiv:1910.13437
-
Big Bidirectional Insertion Representations for Documents 29 Oct 2019 · 0 repositories · arXiv:1910.13034
-
Transformer-Transducer: End-to-End Speech Recognition with Self-Attention 28 Oct 2019 · 1 repository · arXiv:1910.12977
-
An Adaptive and Momental Bound Method for Stochastic Learning 27 Oct 2019 · 2 repositories · arXiv:1910.12249Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Training ASR models by Generation of Contextual Information 27 Oct 2019 · 0 repositories · arXiv:1910.12367
-
HUBERT Untangles BERT to Improve Transfer across NLP Tasks 25 Oct 2019 · 1 repository · arXiv:1910.12647
-
Mockingjay: Unsupervised Speech Representation Learning with Deep Bidirectional Transformer Encoders 25 Oct 2019 · 7 repositories · arXiv:1910.12638Syntology community repositories only · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 1 pointer-only (licence)
-
Towards Online End-to-end Transformer Automatic Speech Recognition 25 Oct 2019 · 0 repositories · arXiv:1910.11871
-
An Empirical Study of Efficient ASR Rescoring with Transformers 24 Oct 2019 · 0 repositories · arXiv:1910.11450
-
ESPnet-TTS: Unified, Reproducible, and Integratable Open Source End-to-End Text-to-Speech Toolkit 24 Oct 2019 · 3 repositories · arXiv:1910.10909Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Promoting the Knowledge of Source Syntax in Transformer NMT Is Not Needed 24 Oct 2019 · 0 repositories · arXiv:1910.11218
-
A Transformer with Interleaved Self-attention and Convolution for Hybrid Acoustic Models 23 Oct 2019 · 1 repository · arXiv:1910.10352
-
Controlling the Output Length of Neural Machine Translation 23 Oct 2019 · 0 repositories · arXiv:1910.10408
-
Deja-vu: Double Feature Presentation and Iterated Loss in Deep Transformer Networks 23 Oct 2019 · 2 repositories · arXiv:1910.10324
-
TCT: A Cross-supervised Learning Method for Multimodal Sequence Representation 23 Oct 2019 · 0 repositories · arXiv:1911.05186
-
Complex Transformer: A Framework for Modeling Complex-Valued Sequence 22 Oct 2019 · 1 repository · arXiv:1910.10202Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Exchangeable deep neural networks for set-to-set matching and learning 22 Oct 2019 · 2 repositories · arXiv:1910.09972
-
Depth-Adaptive Transformer 22 Oct 2019 · 0 repositories · arXiv:1910.10073
-
Improving Transformer-based Speech Recognition Using Unsupervised Pre-training 22 Oct 2019 · 1 repository · arXiv:1910.09932
-
Sequence-to-sequence Singing Synthesis Using the Feed-forward Transformer 22 Oct 2019 · 0 repositories · arXiv:1910.09989
-
Transformer-based Acoustic Modeling for Hybrid Speech Recognition 22 Oct 2019 · 0 repositories · arXiv:1910.09799
-
Learning to Make Generalizable and Diverse Predictions for Retrosynthesis 21 Oct 2019 · 0 repositories · arXiv:1910.09688
-
Transformer-CNN: Fast and Reliable tool for QSAR 21 Oct 2019 · 1 repository · arXiv:1911.06603
-
Personalized Graph Neural Networks with Attention Mechanism for Session-Aware Recommendation 20 Oct 2019 · 3 repositories · arXiv:1910.08887
-
Fully Quantized Transformer for Machine Translation 17 Oct 2019 · 0 repositories · arXiv:1910.10485
-
Predicting retrosynthetic pathways using a combined linguistic model and hyper-graph exploration strategy 17 Oct 2019 · 0 repositories · arXiv:1910.08036
-
Question Classification with Deep Contextualized Transformer 17 Oct 2019 · 1 repository · arXiv:1910.10492
-
Efficiency through Auto-Sizing: Notre Dame NLP's Submission to the WNGT 2019 Efficiency Task 16 Oct 2019 · 0 repositories · arXiv:1910.07134
-
Evolution of transfer learning in natural language processing 16 Oct 2019 · 0 repositories · arXiv:1910.07370
-
Imperial College London Submission to VATEX Video Captioning Task 16 Oct 2019 · 0 repositories · arXiv:1910.07482
-
Injecting Hierarchy with U-Net Transformers 16 Oct 2019 · 2 repositories · arXiv:1910.10488
-
Analyzing the Forgetting Problem in the Pretrain-Finetuning of Dialogue Response Models 16 Oct 2019 · 0 repositories · arXiv:1910.07117
-
Transformer ASR with Contextual Block Processing 16 Oct 2019 · 0 repositories · arXiv:1910.07204
-
Using Whole Document Context in Neural Machine Translation 16 Oct 2019 · 0 repositories · arXiv:1910.07481
-
Enhancing the Transformer with Explicit Relational Encoding for Math Problem Solving 15 Oct 2019 · 3 repositories · arXiv:1910.06611Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples)
-
Facebook AI's WAT19 Myanmar-English Translation Task Submission 15 Oct 2019 · 0 repositories · arXiv:1910.06848
-
Structured Pruning of a BERT-based Question Answering Model 14 Oct 2019 · 0 repositories · arXiv:1910.06360
-
Q8BERT: Quantized 8Bit BERT 14 Oct 2019 · 5 repositories · arXiv:1910.06188Syntology official (archive's flag): 5 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 2 honoured, 0 violated, 7 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 3 pointer-only (licence)
-
Transformers without Tears: Improving the Normalization of Self-Attention 14 Oct 2019 · 5 repositories · arXiv:1910.05895Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Stabilizing Transformers for Reinforcement Learning 13 Oct 2019 · 5 repositories · arXiv:1910.06764Syntology 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
On Recognizing Texts of Arbitrary Shapes with 2D Self-Attention 10 Oct 2019 · 2 repositories · arXiv:1910.04396
-
On the adequacy of untuned warmup for adaptive optimization 9 Oct 2019 · 1 repository · arXiv:1910.04209Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
PipeMare: Asynchronous Pipeline Parallel DNN Training 9 Oct 2019 · 0 repositories · arXiv:1910.05124
-
HuggingFace's Transformers: State-of-the-art Natural Language Processing 9 Oct 2019 · 9 repositories · arXiv:1910.03771Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
SMArT: Training Shallow Memory-aware Transformers for Robotic Explainability 7 Oct 2019 · 0 repositories · arXiv:1910.02974
-
How Transformer Revitalizes Character-based Neural Machine Translation: An Investigation on Japanese-Vietnamese Translation Systems 5 Oct 2019 · 1 repository · arXiv:1910.02238
-
Neural Zero-Inflated Quality Estimation Model For Automatic Speech Recognition System 3 Oct 2019 · 0 repositories · arXiv:1910.01289
-
Application of Low-resource Machine Translation Techniques to Russian-Tatar Language Pair 1 Oct 2019 · 0 repositories · arXiv:1910.00368
-
Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation 1 Oct 2019 · 1 repository · arXiv:1910.06717
-
Dialogue Transformers 1 Oct 2019 · 1 repository · arXiv:1910.00486
-
Efficiency Metrics for Data-Driven Models: A Text Summarization Case Study 1 Oct 2019 · 0 repositories
-
Entangled Transformer for Image Captioning 1 Oct 2019 · 0 repositories
-
Grammatical Error Correction in Low-Resource Scenarios 1 Oct 2019 · 1 repository · arXiv:1910.00353
-
Better Document-Level Machine Translation with Bayes' Rule 1 Oct 2019 · 0 repositories · arXiv:1910.00553
-
Monotonic Multihead Attention 26 Sep 2019 · 3 repositories · arXiv:1909.12406
-
Universal Graph Transformer Self-Attention Networks 26 Sep 2019 · 1 repository · arXiv:1909.11855Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Attention Convolutional Binary Neural Tree for Fine-Grained Visual Categorization 25 Sep 2019 · 2 repositories · arXiv:1909.11378
-
Reducing Transformer Depth on Demand with Structured Dropout 25 Sep 2019 · 5 repositories · arXiv:1909.11556
-
Knowledge-Enriched Transformer for Emotion Detection in Textual Conversations 24 Sep 2019 · 1 repository · arXiv:1909.10681Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Efficiently Reusing Old Models Across Languages via Transfer Learning 24 Sep 2019 · 0 repositories · arXiv:1909.10955
-
Unified Vision-Language Pre-Training for Image Captioning and VQA 24 Sep 2019 · 3 repositories · arXiv:1909.11059Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
TinyBERT: Distilling BERT for Natural Language Understanding 23 Sep 2019 · 10 repositories · arXiv:1909.10351Syntology official: harvested, nothing ran · 0 ran · 4 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Self-attention based end-to-end Hindi-English Neural Machine Translation 21 Sep 2019 · 0 repositories · arXiv:1909.09779
-
Adversarial Learning of General Transformations for Data Augmentation 21 Sep 2019 · 0 repositories · arXiv:1909.09801
-
Improved Variational Neural Machine Translation by Promoting Mutual Information 19 Sep 2019 · 0 repositories · arXiv:1909.09237
-
Language models and Automated Essay Scoring 18 Sep 2019 · 1 repository · arXiv:1909.09482
-
Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism 17 Sep 2019 · 10 repositories · arXiv:1909.08053Syntology community repositories only · 12 ran (of which 2 constructed an object rather than computing a result; 7 with no instrument failure: 4 honoured, 0 violated, 3 with no contract checked; 5 where Syntology's instrument failed) · 35 unverified (of 47 harvested samples) · 15 pointer-only (licence)
-
Hybrid Neural Models For Sequence Modelling: The Best Of Three Worlds 16 Sep 2019 · 0 repositories · arXiv:1909.07102
-
Multilingual Neural Machine Translation for Zero-Resource Languages 16 Sep 2019 · 1 repository · arXiv:1909.07342
-
Automatically Extracting Challenge Sets for Non local Phenomena in Neural Machine Translation 15 Sep 2019 · 1 repository · arXiv:1909.06814