Methods › General › Position Embeddings › Absolute Position Encodings › Papers, page 125
Absolute Position Encodings
Papers archive 2025-07-28
archive papers tagged: 13,942 · with a code link: 6,505 · where Syntology ran a sample: 2,224 (1,897 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,224 of 13,942 tagged: 1,897 with a run with no instrument failure, 327 where every run was a failure of Syntology's instrument)
Page 125 of 140: papers 12,401 to 12,500 of 13,942, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
UPDeT: Universal Multi-agent Reinforcement Learning via Policy Decoupling with Transformers 20 Jan 2021 · 1 repository · arXiv:2101.08001
-
Fast Convergence of DETR with Spatially Modulated Co-Attention 19 Jan 2021 · 2 repositories · arXiv:2101.07448Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Dual-Level Collaborative Transformer for Image Captioning 16 Jan 2021 · 1 repository · arXiv:2101.06462
-
Match-Ignition: Plugging PageRank into Transformer for Long-form Text Matching 16 Jan 2021 · 1 repository · arXiv:2101.06423
-
Exploration of Visual Features and their weighted-additive fusion for Video Captioning 14 Jan 2021 · 0 repositories · arXiv:2101.05806
-
Training Data Leakage Analysis in Language Models 14 Jan 2021 · 0 repositories · arXiv:2101.05405
-
Coarse and Fine-Grained Hostility Detection in Hindi Posts using Fine Tuned Multilingual Embeddings 13 Jan 2021 · 1 repository · arXiv:2101.04998
-
Neural News Recommendation with Negative Feedback 12 Jan 2021 · 0 repositories · arXiv:2101.04328
-
BERT-GT: Cross-sentence n-ary relation extraction with BERT and Graph Transformer 11 Jan 2021 · 0 repositories · arXiv:2101.04158
-
Investigating the Vision Transformer Model for Image Retrieval Tasks 11 Jan 2021 · 0 repositories · arXiv:2101.03771
-
Revisiting Mahalanobis Distance for Transformer-Based Out-of-Domain Detection 11 Jan 2021 · 1 repository · arXiv:2101.03778
-
Spherical Transformer: Adapting Spherical Signal to CNNs 11 Jan 2021 · 0 repositories · arXiv:2101.03848
-
Channel Boosting Feature Ensemble for Radar-based Object Detection 10 Jan 2021 · 0 repositories · arXiv:2101.03531
-
Deep Reinforcement Learning with Function Properties in Mean Reversion Strategies 9 Jan 2021 · 1 repository · arXiv:2101.03418
-
Trankit: A Light-Weight Transformer-based Toolkit for Multilingual Natural Language Processing 9 Jan 2021 · 1 repository · arXiv:2101.03289
-
Leveraging Multilingual Transformers for Hate Speech Detection 8 Jan 2021 · 1 repository · arXiv:2101.03207
-
Compound Word Transformer: Learning to Compose Full-Song Music over Dynamic Directed Hypergraphs 7 Jan 2021 · 5 repositories · arXiv:2101.02402Syntology community repositories only · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 9 harvested samples)
-
TrackFormer: Multi-Object Tracking with Transformers 7 Jan 2021 · 2 repositories · arXiv:2101.02702
-
AutoDropout: Learning Dropout Patterns to Regularize Deep Networks 5 Jan 2021 · 1 repository · arXiv:2101.01761
-
I-BERT: Integer-only BERT Quantization 5 Jan 2021 · 7 repositories · arXiv:2101.01321
-
Improving Portuguese Semantic Role Labeling with Transformers and Transfer Learning 4 Jan 2021 · 1 repository · arXiv:2101.01213
-
Transformers in Vision: A Survey 4 Jan 2021 · 0 repositories · arXiv:2101.01169
-
An Efficient Transformer Decoder with Compressed Sub-layers 3 Jan 2021 · 0 repositories · arXiv:2101.00542
-
Analogical Reasoning for Visually Grounded Compositional Generalization 1 Jan 2021 · 0 repositories
-
AriEL: Volume Coding for Sentence Generation Comparisons 1 Jan 2021 · 0 repositories
-
Attention Is Not Enough: Mitigating the Distribution Discrepancy in Asynchronous Multimodal Sequence Fusion 1 Jan 2021 · 0 repositories
-
Block Skim Transformer for Efficient Question Answering 1 Jan 2021 · 0 repositories
-
Cluster-Former: Clustering-based Sparse Transformer for Question Answering 1 Jan 2021 · 0 repositories
-
CrackFormer: Transformer Network for Fine-Grained Crack Detection 1 Jan 2021 · 1 repository
-
Deep Representational Re-tuning using Contrastive Tension 1 Jan 2021 · 1 repository
-
Discovering Human Interactions With Large-Vocabulary Objects via Query and Multi-Scale Detection 1 Jan 2021 · 0 repositories
-
Do Transformers Understand Polynomial Simplification? 1 Jan 2021 · 0 repositories
-
Dynamic DETR: End-to-End Object Detection With Dynamic Attention 1 Jan 2021 · 0 repositories
-
Event-Based Video Reconstruction Using Transformer 1 Jan 2021 · 1 repository
-
Exploring Routing Strategies for Multilingual Mixture-of-Experts Models 1 Jan 2021 · 0 repositories
-
Frequency-Aware Spatiotemporal Transformers for Video Inpainting Detection 1 Jan 2021 · 0 repositories
-
Generalizing Tree Models for Improving Prediction Accuracy 1 Jan 2021 · 0 repositories
-
High-Performance Discriminative Tracking With Transformers 1 Jan 2021 · 0 repositories
-
HyperGrid Transformers: Towards A Single Model for Multiple Tasks 1 Jan 2021 · 0 repositories
-
Image Harmonization With Transformer 1 Jan 2021 · 1 repository
-
Improving Generalizability of Protein Sequence Models via Data Augmentations 1 Jan 2021 · 0 repositories
-
Improving Machine Translation by Searching Skip Connections Efficiently 1 Jan 2021 · 0 repositories
-
KETG: A Knowledge Enhanced Text Generation Framework 1 Jan 2021 · 0 repositories
-
Long Range Arena : A Benchmark for Efficient Transformers 1 Jan 2021 · 0 repositories
-
Memory Representation in Transformer 1 Jan 2021 · 0 repositories
-
Multi-View 3D Reconstruction With Transformers 1 Jan 2021 · 0 repositories
-
Multimodal Co-Attention Transformer for Survival Prediction in Gigapixel Whole Slide Images 1 Jan 2021 · 1 repository
-
Non-iterative Parallel Text Generation via Glancing Transformer 1 Jan 2021 · 0 repositories
-
On Position Embeddings in BERT 1 Jan 2021 · 0 repositories
-
Parameterization of Hypercomplex Multiplications 1 Jan 2021 · 0 repositories
-
PhraseTransformer: Self-Attention using Local Context for Semantic Parsing 1 Jan 2021 · 1 repository
-
Post-Training Weighted Quantization of Neural Networks for Language Models 1 Jan 2021 · 0 repositories
-
Predictive Attention Transformer: Improving Transformer with Attention Map Prediction 1 Jan 2021 · 0 repositories
-
Representation and Bias in Multilingual NLP: Insights from Controlled Experiments on Conditional Language Modeling 1 Jan 2021 · 0 repositories
-
Representational correlates of hierarchical phrase structure in deep language models 1 Jan 2021 · 0 repositories
-
Scene Context-Aware Salient Object Detection 1 Jan 2021 · 1 repository
-
Share or Not? Learning to Schedule Language-Specific Capacity for Multilingual Translation 1 Jan 2021 · 0 repositories
-
Single Layers of Attention Suffice to Predict Protein Contacts 1 Jan 2021 · 0 repositories
-
STAR: A Structure-Aware Lightweight Transformer for Real-Time Image Enhancement 1 Jan 2021 · 1 repository
-
Subformer: A Parameter Reduced Transformer 1 Jan 2021 · 0 repositories
-
Subformer: Exploring Weight Sharing for Parameter Efficiency in Generative Transformers 1 Jan 2021 · 1 repository · arXiv:2101.00234
-
Synthesizer: Rethinking Self-Attention for Transformer Models 1 Jan 2021 · 0 repositories
-
Trans-Caps: Transformer Capsule Networks with Self-attention Routing 1 Jan 2021 · 0 repositories
-
Transformer protein language models are unsupervised structure learners 1 Jan 2021 · 0 repositories
-
Transformer-QL: A Step Towards Making Transformer Network Quadratically Large 1 Jan 2021 · 0 repositories
-
Transformers satisfy 1 Jan 2021 · 0 repositories
-
Transforming Recurrent Neural Networks with Attention and Fixed-point Equations 1 Jan 2021 · 0 repositories
-
TRAR: Routing the Attention Spans in Transformer for Visual Question Answering 1 Jan 2021 · 1 repository
-
UPDeT: Universal Multi-agent RL via Policy Decoupling with Transformers 1 Jan 2021 · 0 repositories
-
Visual Transformers: Where Do Transformers Really Belong in Vision Models? 1 Jan 2021 · 0 repositories
-
VisualSparta: An Embarrassingly Simple Approach to Large-scale Text-to-Image Search with Weighted Bag-of-words 1 Jan 2021 · 1 repository · arXiv:2101.00265
-
WB-DETR: Transformer-Based Detector Without Backbone 1 Jan 2021 · 0 repositories
-
A Multi-modal Deep Learning Model for Video Thumbnail Selection 31 Dec 2020 · 0 repositories · arXiv:2101.00073
-
Fully Non-autoregressive Neural Machine Translation: Tricks of the Trade 31 Dec 2020 · 1 repository · arXiv:2012.15833
-
MiniLMv2: Multi-Head Self-Attention Relation Distillation for Compressing Pretrained Transformers 31 Dec 2020 · 2 repositories · arXiv:2012.15828
-
Revisiting Robust Neural Machine Translation: A Transformer Case Study 31 Dec 2020 · 0 repositories · arXiv:2012.15710
-
TransTrack: Multiple Object Tracking with Transformer 31 Dec 2020 · 2 repositories · arXiv:2012.15460
-
Verb Knowledge Injection for Multilingual Event Processing 31 Dec 2020 · 0 repositories · arXiv:2012.15421
-
XLM-T: Scaling up Multilingual Machine Translation with Pretrained Cross-lingual Transformer Encoders 31 Dec 2020 · 0 repositories · arXiv:2012.15547
-
Optimizing Deeper Transformers on Small Datasets 30 Dec 2020 · 1 repository · arXiv:2012.15355
-
Transformer for Image Quality Assessment 30 Dec 2020 · 0 repositories · arXiv:2101.01097
-
UnNatural Language Inference 30 Dec 2020 · 1 repository · arXiv:2101.00010
-
A Hierarchical Transformer with Speaker Modeling for Emotion Recognition in Conversation 29 Dec 2020 · 1 repository · arXiv:2012.14781
-
Kaleidoscope: An Efficient, Learnable Representation For All Structured Linear Maps 29 Dec 2020 · 2 repositories · arXiv:2012.14966Syntology official (archive's flag): 16 ran · 16 ran (of which 9 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 7 where Syntology's instrument failed) · 7 unverified (of 23 harvested samples)
-
LayoutLMv2: Multi-modal Pre-training for Visually-Rich Document Understanding 29 Dec 2020 · 9 repositories · arXiv:2012.14740
-
Code Summarization with Structure-induced Transformer 29 Dec 2020 · 1 repository · arXiv:2012.14710
-
Lattice-Free MMI Adaptation Of Self-Supervised Pretrained Acoustic Models 28 Dec 2020 · 2 repositories · arXiv:2012.14252
-
Red Dragon AI at TextGraphs 2020 Shared Task: LIT : LSTM-Interleaved Transformer for Multi-Hop Explanation Ranking 28 Dec 2020 · 1 repository · arXiv:2012.14164
-
Syntax-Enhanced Pre-trained Model 28 Dec 2020 · 1 repository · arXiv:2012.14116
-
TransPose: Keypoint Localization via Transformer 28 Dec 2020 · 1 repository · arXiv:2012.14214
-
Learning Light-Weight Translation Models from Deep Transformer 27 Dec 2020 · 1 repository · arXiv:2012.13866Syntology official (archive's flag): 1 ran · 2 ran (of which 2 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples) · 1 pointer-only (licence)
-
Portfolio Optimization with 2D Relative-Attentional Gated Transformer 27 Dec 2020 · 0 repositories · arXiv:2101.03138
-
SG-Net: Syntax Guided Transformer for Language Representation 27 Dec 2020 · 0 repositories · arXiv:2012.13915
-
Detecting Hateful Memes Using a Multimodal Deep Ensemble 24 Dec 2020 · 1 repository · arXiv:2012.13235
-
I like fish, especially dolphins: Addressing Contradictions in Dialogue Modeling 24 Dec 2020 · 0 repositories · arXiv:2012.13391
-
Future-Guided Incremental Transformer for Simultaneous Translation 23 Dec 2020 · 0 repositories · arXiv:2012.12465
-
Domain Adaptation of NMT models for English-Hindi Machine Translation Task at AdapMT ICON 2020 22 Dec 2020 · 0 repositories · arXiv:2012.12112
-
Molecular CT: Unifying Geometry and Representation Learning for Molecules at Different Scales 22 Dec 2020 · 0 repositories · arXiv:2012.11816
-
Multi-Head Self-Attention with Role-Guided Masks 22 Dec 2020 · 1 repository · arXiv:2012.12366
-
3D Object Detection with Pointformer 21 Dec 2020 · 1 repository · arXiv:2012.11409