Methods › General › Attention Mechanisms › Attention › Papers, page 243
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 243 of 316: papers 24,201 to 24,300 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Transformers for 1D Signals in Parkinson's Disease Detection from Gait 1 Apr 2022 · 1 repository · arXiv:2204.00423
-
Vision Transformer with Cross-attention by Temporal Shift for Efficient Action Recognition 1 Apr 2022 · 0 repositories · arXiv:2204.00452
-
A Baseline Readability Model for Cebuano 31 Mar 2022 · 1 repository · arXiv:2203.17225
-
A Character-level Span-based Model for Mandarin Prosodic Structure Prediction 31 Mar 2022 · 1 repository · arXiv:2203.16922
-
An End-to-end Chinese Text Normalization Model based on Rule-guided Flat-Lattice Transformer 31 Mar 2022 · 1 repository · arXiv:2203.16954
-
CatIss: An Intelligent Tool for Categorizing Issues Reports using Transformers 31 Mar 2022 · 1 repository · arXiv:2203.17196
-
CRAFT: Cross-Attentional Flow Transformer for Robust Optical Flow 31 Mar 2022 · 1 repository · arXiv:2203.16896Syntology official (archive's flag): 7 ran · 7 ran (of which 4 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
CREATE: A Benchmark for Chinese Short Video Retrieval and Title Generation 31 Mar 2022 · 0 repositories · arXiv:2203.16763
-
Deformable Video Transformer 31 Mar 2022 · 0 repositories · arXiv:2203.16795
-
ESGBERT: Language Model to Help with Classification Tasks Related to Companies Environmental, Social, and Governance Practices 31 Mar 2022 · 0 repositories · arXiv:2203.16788
-
Generative Pre-Trained Transformers for Biologically Inspired Design 31 Mar 2022 · 0 repositories · arXiv:2204.09714
-
Leveraging pre-trained language models for conversational information seeking from text 31 Mar 2022 · 0 repositories · arXiv:2204.03542
-
Mixed-Phoneme BERT: Improving BERT with Mixed Phoneme and Sup-Phoneme Representations for Text to Speech 31 Mar 2022 · 0 repositories · arXiv:2203.17190
-
Scaling Language Model Size in Cross-Device Federated Learning 31 Mar 2022 · 0 repositories · arXiv:2204.09715
-
Universal Lymph Node Detection in T2 MRI using Neural Networks 31 Mar 2022 · 0 repositories · arXiv:2204.00622
-
AdaMixer: A Fast-Converging Query-Based Object Detector 30 Mar 2022 · 2 repositories · arXiv:2203.16507Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
An Efficient Anchor-free Universal Lesion Detection in CT-scans 30 Mar 2022 · 0 repositories · arXiv:2203.16074
-
Collaborative Transformers for Grounded Situation Recognition 30 Mar 2022 · 3 repositories · arXiv:2203.16518Syntology official (archive's flag): 2 ran · 5 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
End-to-End Table Question Answering via Retrieval-Augmented Generation 30 Mar 2022 · 0 repositories · arXiv:2203.16714
-
Exploring Plain Vision Transformer Backbones for Object Detection 30 Mar 2022 · 11 repositories · arXiv:2203.16527Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Incorporating Dynamic Semantics into Pre-Trained Language Model for Aspect-based Sentiment Analysis 30 Mar 2022 · 0 repositories · arXiv:2203.16369
-
ITTR: Unpaired Image-to-Image Translation with Transformers 30 Mar 2022 · 0 repositories · arXiv:2203.16015
-
Speculative Decoding: Exploiting Speculative Execution for Accelerating Seq2seq Generation 30 Mar 2022 · 2 repositories · arXiv:2203.16487Syntology official (archive's flag): 1 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
MAE-AST: Masked Autoencoding Audio Spectrogram Transformer 30 Mar 2022 · 2 repositories · arXiv:2203.16691
-
Spatial-Temporal Parallel Transformer for Arm-Hand Dynamic Estimation 30 Mar 2022 · 0 repositories · arXiv:2203.16202
-
Surface Vision Transformers: Attention-Based Modelling applied to Cortical Analysis 30 Mar 2022 · 1 repository · arXiv:2203.16414
-
Tampered VAE for Improved Satellite Image Time Series Classification 30 Mar 2022 · 0 repositories · arXiv:2203.16149
-
Transformer Language Models without Positional Encodings Still Learn Positional Information 30 Mar 2022 · 1 repository · arXiv:2203.16634Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
TubeDETR: Spatio-Temporal Video Grounding with Transformers 30 Mar 2022 · 1 repository · arXiv:2203.16434Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 3 harvested samples)
-
A Fast Post-Training Pruning Framework for Transformers 29 Mar 2022 · 2 repositories · arXiv:2204.09656Syntology official (archive's flag): 1 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Affine Medical Image Registration with Coarse-to-Fine Vision Transformer 29 Mar 2022 · 1 repository · arXiv:2203.15216Syntology official (archive's flag): 5 ran · 5 ran (of which 4 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
AnoDFDNet: A Deep Feature Difference Network for Anomaly Detection 29 Mar 2022 · 1 repository · arXiv:2203.15195
-
CAT-Net: A Cross-Slice Attention Transformer Model for Prostate Zonal Segmentation in MRI 29 Mar 2022 · 1 repository · arXiv:2203.15163
-
Cross-Modality High-Frequency Transformer for MR Image Super-Resolution 29 Mar 2022 · 0 repositories · arXiv:2203.15314
-
Dynamic Latency for CTC-Based Streaming Automatic Speech Recognition With Emformer 29 Mar 2022 · 0 repositories · arXiv:2203.15613
-
End-to-End Transformer Based Model for Image Captioning 29 Mar 2022 · 2 repositories · arXiv:2203.15350Syntology official (archive's flag): 8 ran · 17 ran (of which 9 constructed an object rather than computing a result; 14 with no instrument failure: 1 honoured, 2 violated, 11 with no contract checked; 3 where Syntology's instrument failed) · 16 unverified (of 33 harvested samples) · 15 pointer-only (licence)
-
Exploring Intra- and Inter-Video Relation for Surgical Semantic Scene Segmentation 29 Mar 2022 · 1 repository · arXiv:2203.15251
-
Fine-tuning Image Transformers using Learnable Memory 29 Mar 2022 · 1 repository · arXiv:2203.15243Syntology 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Hybrid Routing Transformer for Zero-Shot Learning 29 Mar 2022 · 0 repositories · arXiv:2203.15310
-
Improving Persian Relation Extraction Models by Data Augmentation 29 Mar 2022 · 0 repositories · arXiv:2203.15323
-
In-N-Out Generative Learning for Dense Unsupervised Video Segmentation 29 Mar 2022 · 1 repository · arXiv:2203.15312
-
Integrative Few-Shot Learning for Classification and Segmentation 29 Mar 2022 · 1 repository · arXiv:2203.15712
-
LightHuBERT: Lightweight and Configurable Speech Representation Learning with Once-for-All Hidden-Unit BERT 29 Mar 2022 · 1 repository · arXiv:2203.15610Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
LinkBERT: Pretraining Language Models with Document Links 29 Mar 2022 · 1 repository · arXiv:2203.15827Syntology official: harvested, nothing ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 11 unverified (of 14 harvested samples)
-
MatteFormer: Transformer-Based Image Matting via Prior-Tokens 29 Mar 2022 · 1 repository · arXiv:2203.15662Syntology official (archive's flag): 4 ran · 5 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
mc-BEiT: Multi-choice Discretization for Image BERT Pre-training 29 Mar 2022 · 1 repository · arXiv:2203.15371Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
Protein language models trained on multiple sequence alignments learn phylogenetic relationships 29 Mar 2022 · 1 repository · arXiv:2203.15465
-
SepViT: Separable Vision Transformer 29 Mar 2022 · 2 repositories · arXiv:2203.15380Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Shifted Chunk Encoder for Transformer Based Streaming End-to-End ASR 29 Mar 2022 · 1 repository · arXiv:2203.15206
-
Training Compute-Optimal Large Language Models 29 Mar 2022 · 2 repositories · arXiv:2203.15556Syntology 8 ran (of which 3 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 4 pointer-only (licence)
-
Transformer Inertial Poser: Real-time Human Motion Reconstruction from Sparse IMUs with Simultaneous Terrain Generation 29 Mar 2022 · 1 repository · arXiv:2203.15720Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Treatment Learning Causal Transformer for Noisy Image Classification 29 Mar 2022 · 0 repositories · arXiv:2203.15529
-
Unified Transformer Tracker for Object Tracking 29 Mar 2022 · 1 repository · arXiv:2203.15175Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
VPTR: Efficient Transformers for Video Prediction 29 Mar 2022 · 1 repository · arXiv:2203.15836Syntology official: harvested, nothing ran · 0 ran · 7 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
ANNA: Enhanced Language Representation for Question Answering 28 Mar 2022 · 0 repositories · arXiv:2203.14507
-
Automated Progressive Learning for Efficient Training of Vision Transformers 28 Mar 2022 · 2 repositories · arXiv:2203.14509Syntology official (archive's flag): 7 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 13 harvested samples) · 4 pointer-only (licence)
-
Continuous Metric Learning For Transferable Speech Emotion Recognition and Embedding Across Low-resource Languages 28 Mar 2022 · 0 repositories · arXiv:2203.14867
-
Discovering material information using hierarchical Reformer model on financial regulatory filings 28 Mar 2022 · 0 repositories · arXiv:2204.05979
-
Diverse Plausible 360-Degree Image Outpainting for Efficient 3DCG Background Creation 28 Mar 2022 · 1 repository · arXiv:2203.14668Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Enhancing Neural Mathematical Reasoning by Abductive Combination with Symbolic Library 28 Mar 2022 · 0 repositories · arXiv:2203.14487
-
Hierarchical Transformer Model for Scientific Named Entity Recognition 28 Mar 2022 · 1 repository · arXiv:2203.14710
-
Integrating Physiological Time Series and Clinical Notes with Transformer for Early Prediction of Sepsis 28 Mar 2022 · 0 repositories · arXiv:2203.14469
-
Curriculum learning for self-supervised speaker verification 28 Mar 2022 · 0 repositories · arXiv:2203.14525
-
Stratified Transformer for 3D Point Cloud Segmentation 28 Mar 2022 · 4 repositories · arXiv:2203.14508Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 5 pointer-only (licence)
-
TraHGR: Transformer for Hand Gesture Recognition via ElectroMyography 28 Mar 2022 · 0 repositories · arXiv:2203.16336
-
UTSA NLP at SemEval-2022 Task 4: An Exploration of Simple Ensembles of Transformers, Convolutional, and Recurrent Neural Networks 28 Mar 2022 · 0 repositories · arXiv:2203.14920
-
Visual Mechanisms Inspired Efficient Transformers for Image and Video Quality Assessment 28 Mar 2022 · 0 repositories · arXiv:2203.14557
-
BARCOR: Towards A Unified Framework for Conversational Recommendation Systems 27 Mar 2022 · 0 repositories · arXiv:2203.14257
-
DepthFormer: Exploiting Long-Range Correlation and Local Information for Accurate Monocular Depth Estimation 27 Mar 2022 · 1 repository · arXiv:2203.14211
-
Error Correction Code Transformer 27 Mar 2022 · 2 repositories · arXiv:2203.14966Syntology official (archive's flag): 1 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Example-based Hypernetworks for Out-of-Distribution Generalization 27 Mar 2022 · 1 repository · arXiv:2203.14276
-
Leveraging Search History for Improving Person-Job Fit 27 Mar 2022 · 0 repositories · arXiv:2203.14232
-
Pyramid-BERT: Reducing Complexity via Successive Core-set based Token Selection 27 Mar 2022 · 0 repositories · arXiv:2203.14380
-
RSTT: Real-time Spatial Temporal Transformer for Space-Time Video Super-Resolution 27 Mar 2022 · 1 repository · arXiv:2203.14186Syntology official (archive's flag): 8 ran · 8 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 6 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
StruBERT: Structure-aware BERT for Table Search and Matching 27 Mar 2022 · 1 repository · arXiv:2203.14278Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples)
-
Autoregressive Linguistic Steganography Based on BERT and Consistency Coding 26 Mar 2022 · 0 repositories · arXiv:2203.13972
-
Feature Selective Transformer for Semantic Image Segmentation 26 Mar 2022 · 0 repositories · arXiv:2203.14124
-
3D-OAE: Occlusion Auto-Encoders for Self-Supervised Learning on Point Clouds 26 Mar 2022 · 1 repository · arXiv:2203.14084
-
Semantic Segmentation by Early Region Proxy 26 Mar 2022 · 1 repository · arXiv:2203.14043Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Give Me Your Attention: Dot-Product Attention Considered Harmful for Adversarial Patch Robustness 25 Mar 2022 · 0 repositories · arXiv:2203.13639
-
GPT-D: Inducing Dementia-related Linguistic Anomalies by Deliberate Degradation of Artificial Neural Language Models 25 Mar 2022 · 2 repositories · arXiv:2203.13397
-
Gransformer: Transformer-based Graph Generation 25 Mar 2022 · 0 repositories · arXiv:2203.13655
-
High-Performance Transformer Tracking 25 Mar 2022 · 1 repository · arXiv:2203.13533
-
Implicit Neural Representations for Variable Length Human Motion Generation 25 Mar 2022 · 1 repository · arXiv:2203.13694Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
L3Cube-MahaHate: A Tweet-based Marathi Hate Speech Detection Dataset and BERT models 25 Mar 2022 · 1 repository · arXiv:2203.13778
-
MKQ-BERT: Quantized BERT with 4-bits Weights and Activations 25 Mar 2022 · 0 repositories · arXiv:2203.13483
-
Modeling Target-Side Morphology in Neural Machine Translation: A Comparison of Strategies 25 Mar 2022 · 0 repositories · arXiv:2203.13550
-
Predicting Clinical Intent from Free Text Electronic Health Records 25 Mar 2022 · 0 repositories · arXiv:2204.09594
-
Vision Transformer Compression with Structured Pruning and Low Rank Approximation 25 Mar 2022 · 0 repositories · arXiv:2203.13444
-
An Ensemble Approach for Facial Expression Analysis in Video 24 Mar 2022 · 0 repositories · arXiv:2203.12891
-
Bailando: 3D Dance Generation by Actor-Critic GPT with Choreographic Memory 24 Mar 2022 · 1 repository · arXiv:2203.13055Syntology official (archive's flag): 4 ran · 4 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Beyond Fixation: Dynamic Window Visual Transformer 24 Mar 2022 · 1 repository · arXiv:2203.12856
-
CrossFormer: Cross Spatio-Temporal Transformer for 3D Human Pose Estimation 24 Mar 2022 · 1 repository · arXiv:2203.13387
-
Deep learning for laboratory earthquake prediction and autoregressive forecasting of fault zone stress 24 Mar 2022 · 0 repositories · arXiv:2203.13313
-
Ensembling and Knowledge Distilling of Large Sequence Taggers for Grammatical Error Correction 24 Mar 2022 · 1 repository · arXiv:2203.13064
-
Evaluating Distributional Distortion in Neural Language Modeling 24 Mar 2022 · 0 repositories · arXiv:2203.12788
-
Facial Expression Classification using Fusion of Deep Neural Network in Video for the 3rd ABAW3 Competition 24 Mar 2022 · 0 repositories · arXiv:2203.12899
-
Industrial Style Transfer with Large-scale Geometric Warping and Content Preservation 24 Mar 2022 · 1 repository · arXiv:2203.12835
-
mcBERT: Momentum Contrastive Learning with BERT for Zero-Shot Slot Filling 24 Mar 2022 · 0 repositories · arXiv:2203.12940
-
minicons: Enabling Flexible Behavioral and Representational Analyses of Transformer Language Models 24 Mar 2022 · 1 repository · arXiv:2203.13112