Methods › General › Attention Mechanisms › Attention › Papers, page 248
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 248 of 316: papers 24,701 to 24,800 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Robust Training of Neural Networks Using Scale Invariant Architectures 2 Feb 2022 · 0 repositories · arXiv:2202.00980
-
An Adaptive Deep Clustering Pipeline to Inform Text Labeling at Scale 1 Feb 2022 · 0 repositories · arXiv:2202.01211
-
A Semi-Supervised Deep Clustering Pipeline for Mining Intentions From Texts 1 Feb 2022 · 0 repositories · arXiv:2202.00802
-
AlphaDesign: A graph protein design method and benchmark on AlphaFoldDB 1 Feb 2022 · 1 repository · arXiv:2202.01079
-
LayoutEnhancer: Generating Good Indoor Layouts from Imperfect Data 1 Feb 2022 · 1 repository · arXiv:2202.00185Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples)
-
Improving BERT-based Query-by-Document Retrieval with Multi-Task Optimization 1 Feb 2022 · 0 repositories · arXiv:2202.00373
-
Is the Performance of My Deep Network Too Good to Be True? A Direct Approach to Estimating the Bayes Error in Binary Classification 1 Feb 2022 · 1 repository · arXiv:2202.00395Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Local Feature Matching with Transformers for low-end devices 1 Feb 2022 · 1 repository · arXiv:2202.00770
-
Regression Transformer: Concurrent sequence regression and generation for molecular language modeling 1 Feb 2022 · 1 repository · arXiv:2202.01338Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
Transformer-based Models of Text Normalization for Speech Applications 1 Feb 2022 · 0 repositories · arXiv:2202.00153
-
BOAT: Bilateral Local Attention Vision Transformer 31 Jan 2022 · 1 repository · arXiv:2201.13027
-
Memory-Efficient Backpropagation through Large Linear Layers 31 Jan 2022 · 2 repositories · arXiv:2201.13195
-
Fast Monte-Carlo Approximation of the Attention Mechanism 30 Jan 2022 · 0 repositories · arXiv:2201.12854
-
FEDformer: Frequency Enhanced Decomposed Transformer for Long-term Series Forecasting 30 Jan 2022 · 3 repositories · arXiv:2201.12740Syntology community repositories only · 14 ran (of which 8 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 18 harvested samples)
-
GRPE: Relative Positional Encoding for Graph Transformer 30 Jan 2022 · 1 repository · arXiv:2201.12787Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
N-HiTS: Neural Hierarchical Interpolation for Time Series Forecasting 30 Jan 2022 · 4 repositories · arXiv:2201.12886Syntology community repositories only · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
TransBTSV2: Towards Better and More Efficient Volumetric Segmentation of Medical Images 30 Jan 2022 · 2 repositories · arXiv:2201.12785
-
A Frustratingly Simple Approach for End-to-End Image Captioning 30 Jan 2022 · 0 repositories · arXiv:2201.12723
-
AutoDistil: Few-shot Task-agnostic Neural Architecture Search for Distilling Large Language Models 29 Jan 2022 · 0 repositories · arXiv:2201.12507
-
Decepticons: Corrupted Transformers Breach Privacy in Federated Learning for Language Models 29 Jan 2022 · 1 repository · arXiv:2201.12675
-
Does Transliteration Help Multilingual Language Modeling? 29 Jan 2022 · 1 repository · arXiv:2201.12501Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Research on Patch Attentive Neural Process 29 Jan 2022 · 0 repositories · arXiv:2202.01884
-
Rewiring with Positional Encodings for Graph Neural Networks 29 Jan 2022 · 0 repositories · arXiv:2201.12674
-
Benchmarking Robustness of 3D Point Cloud Recognition Against Common Corruptions 28 Jan 2022 · 6 repositories · arXiv:2201.12296Syntology official (archive's flag): 11 ran · 24 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 2 honoured, 0 violated, 14 with no contract checked; 8 where Syntology's instrument failed) · 5 unverified (of 29 harvested samples) · 7 pointer-only (licence)
-
Can Wikipedia Help Offline Reinforcement Learning? 28 Jan 2022 · 1 repository · arXiv:2201.12122Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Chain-of-Thought Prompting Elicits Reasoning in Large Language Models 28 Jan 2022 · 19 repositories · arXiv:2201.11903Syntology 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
DAB-DETR: Dynamic Anchor Boxes are Better Queries for DETR 28 Jan 2022 · 8 repositories · arXiv:2201.12329Syntology community repositories only · 6 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
Electra: Conditional Generative Model based Predicate-Aware Query Approximation 28 Jan 2022 · 0 repositories · arXiv:2201.12420
-
O-ViT: Orthogonal Vision Transformer 28 Jan 2022 · 0 repositories · arXiv:2201.12133
-
Describing Differences between Text Distributions with Natural Language 28 Jan 2022 · 1 repository · arXiv:2201.12323Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
VRT: A Video Restoration Transformer 28 Jan 2022 · 1 repository · arXiv:2201.12288Syntology official (archive's flag): 3 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Clinical-Longformer and Clinical-BigBird: Transformers for long clinical sequences 27 Jan 2022 · 1 repository · arXiv:2201.11838
-
Generalised Image Outpainting with U-Transformer 27 Jan 2022 · 1 repository · arXiv:2201.11403
-
Going Extreme: Comparative Analysis of Hate Speech in Parler and Gab 27 Jan 2022 · 1 repository · arXiv:2201.11770
-
Grad2Task: Improved Few-shot Text Classification Using Gradients for Task Representation 27 Jan 2022 · 1 repository · arXiv:2201.11576
-
Reasoning Like Program Executors 27 Jan 2022 · 1 repository · arXiv:2201.11473
-
Transformer Module Networks for Systematic Generalization in Visual Question Answering 27 Jan 2022 · 1 repository · arXiv:2201.11316
-
A Comprehensive Study of Image Classification Model Sensitivity to Foregrounds, Backgrounds, and Visual Attributes 26 Jan 2022 · 1 repository · arXiv:2201.10766Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
DiscoScore: Evaluating Text Generation with BERT and Discourse Coherence 26 Jan 2022 · 1 repository · arXiv:2201.11176
-
DNNFuser: Generative Pre-Trained Transformer as a Generalized Mapper for Layer Fusion in DNN Accelerators 26 Jan 2022 · 0 repositories · arXiv:2201.11218
-
DSFormer: A Dual-domain Self-supervised Transformer for Accelerated Multi-contrast MRI Reconstruction 26 Jan 2022 · 0 repositories · arXiv:2201.10776
-
On the Effectiveness of Pinyin-Character Dual-Decoding for End-to-End Mandarin Chinese ASR 26 Jan 2022 · 0 repositories · arXiv:2201.10792
-
Dual-Tasks Siamese Transformer Framework for Building Damage Assessment 26 Jan 2022 · 0 repositories · arXiv:2201.10953
-
FiNCAT: Financial Numeral Claim Analysis Tool 26 Jan 2022 · 1 repository · arXiv:2202.00631
-
Neural Grapheme-to-Phoneme Conversion with Pre-trained Grapheme Models 26 Jan 2022 · 1 repository · arXiv:2201.10716
-
Predicting Knee Osteoarthritis Progression from Structural MRI using Deep Learning 26 Jan 2022 · 1 repository · arXiv:2201.10849
-
Self-supervised 3D Semantic Representation Learning for Vision-and-Language Navigation 26 Jan 2022 · 0 repositories · arXiv:2201.10788
-
Synchromesh: Reliable code generation from pre-trained language models 26 Jan 2022 · 2 repositories · arXiv:2201.11227Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
TransPPG: Two-stream Transformer for Remote Heart Rate Estimate 26 Jan 2022 · 0 repositories · arXiv:2201.10873
-
When Shift Operation Meets Vision Transformer: An Extremely Simple Alternative to Attention Mechanism 26 Jan 2022 · 2 repositories · arXiv:2201.10801Syntology official (archive's flag): 6 ran · 6 ran (of which 6 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 6 samples that ran constructed an object rather than computing a result (of 7 harvested samples)
-
BERTHA: Video Captioning Evaluation Via Transfer-Learned Human Assessment 25 Jan 2022 · 1 repository · arXiv:2201.10243
-
Convolutional Xformers for Vision 25 Jan 2022 · 1 repository · arXiv:2201.10271
-
Explore-And-Match: Bridging Proposal-Based and Proposal-Free With Transformer for Sentence Grounding in Videos 25 Jan 2022 · 1 repository · arXiv:2201.10168
-
Improving non-autoregressive end-to-end speech recognition with pre-trained acoustic and language models 25 Jan 2022 · 0 repositories · arXiv:2201.10103
-
Neighbour Interaction based Click-Through Rate Prediction via Graph-masked Transformer 25 Jan 2022 · 0 repositories · arXiv:2201.13311
-
Cardiac Disease Diagnosis on Imbalanced Electrocardiography Data Through Optimal Transport Augmentation 25 Jan 2022 · 0 repositories · arXiv:2202.00567
-
Pre-Trained Language Transformers are Universal Image Classifiers 25 Jan 2022 · 0 repositories · arXiv:2201.10182
-
ViT-HGR: Vision Transformer-based Hand Gesture Recognition from High Density Surface EMG Signals 25 Jan 2022 · 1 repository · arXiv:2201.10060
-
Whose Language Counts as High Quality? Measuring Language Ideologies in Text Data Selection 25 Jan 2022 · 0 repositories · arXiv:2201.10474
-
Zero-Shot Sketch Based Image Retrieval using Graph Transformer 25 Jan 2022 · 0 repositories · arXiv:2201.10185
-
Emotion-based Modeling of Mental Disorders on Social Media 24 Jan 2022 · 0 repositories · arXiv:2201.09451
-
Improving Chest X-Ray Report Generation by Leveraging Warm Starting 24 Jan 2022 · 1 repository · arXiv:2201.09405
-
Patches Are All You Need? 24 Jan 2022 · 12 repositories · arXiv:2201.09792Syntology official (archive's flag): 6 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 5 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples)
-
Polyphone disambiguation and accent prediction using pre-trained language models in Japanese TTS front-end 24 Jan 2022 · 0 repositories · arXiv:2201.09427
-
Synthetic Books 24 Jan 2022 · 0 repositories · arXiv:2201.09518
-
Transformers in Medical Imaging: A Survey 24 Jan 2022 · 1 repository · arXiv:2201.09873
-
Unified Multimodal Punctuation Restoration Framework for Mixed-Modality Corpus 24 Jan 2022 · 1 repository · arXiv:2202.00468
-
A Large and Diverse Arabic Corpus for Language Modeling 23 Jan 2022 · 0 repositories · arXiv:2201.09227
-
A Pre-trained Audio-Visual Transformer for Emotion Recognition 23 Jan 2022 · 0 repositories · arXiv:2201.09165
-
A Survey for Deep RGBT Tracking 23 Jan 2022 · 0 repositories · arXiv:2201.09296
-
An Application of Pseudo-Log-Likelihoods to Natural Language Scoring 23 Jan 2022 · 0 repositories · arXiv:2201.09377
-
Fast MRI Reconstruction: How Powerful Transformers Are? 23 Jan 2022 · 1 repository · arXiv:2201.09400
-
How Expressive are Transformers in Spectral Domain for Graphs? 23 Jan 2022 · 1 repository · arXiv:2201.09332Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
ReconFormer: Accelerated MRI Reconstruction Using Recurrent Transformer 23 Jan 2022 · 1 repository · arXiv:2201.09376
-
Dual-Flattening Transformers through Decomposed Row and Column Queries for Semantic Segmentation 22 Jan 2022 · 0 repositories · arXiv:2201.09139
-
A Comparative Study on Language Models for Task-Oriented Dialogue Systems 21 Jan 2022 · 1 repository · arXiv:2201.08687
-
AutoDistill: an End-to-End Framework to Explore and Distill Hardware-Efficient Language Models 21 Jan 2022 · 0 repositories · arXiv:2201.08539
-
Black-box Prompt Learning for Pre-trained Language Models 21 Jan 2022 · 1 repository · arXiv:2201.08531Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Dual Contrastive Learning: Text Classification via Label-Aware Data Augmentation 21 Jan 2022 · 2 repositories · arXiv:2201.08702
-
Fast Differentiable Matrix Square Root 21 Jan 2022 · 1 repository · arXiv:2201.08663Syntology official: harvested for another paper · 0 ran · 1 unverified (of 1 harvested sample)
-
Less is Less: When Are Snippets Insufficient for Human vs Machine Relevance Estimation? 21 Jan 2022 · 0 repositories · arXiv:2201.08721
-
Representing Long-Range Context for Graph Neural Networks with Global Attention 21 Jan 2022 · 1 repository · arXiv:2201.08821Syntology official (archive's flag): 7 ran · 7 ran (of which 5 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 7 unverified (of 14 harvested samples)
-
Uncertainty-Cognizant Model Predictive Control for Energy Management of Residential Buildings with PVT and Thermal Energy Storage 21 Jan 2022 · 0 repositories · arXiv:2201.08909
-
Cheating Automatic Short Answer Grading: On the Adversarial Usage of Adjectives and Adverbs 20 Jan 2022 · 1 repository · arXiv:2201.08318
-
MeMViT: Memory-Augmented Multiscale Vision Transformer for Efficient Long-Term Video Recognition 20 Jan 2022 · 1 repository · arXiv:2201.08383
-
Meta Learning for Code Summarization 20 Jan 2022 · 0 repositories · arXiv:2201.08310
-
Sentiment Analysis: Predicting Yelp Scores 20 Jan 2022 · 0 repositories · arXiv:2201.07999
-
TerViT: An Efficient Ternary Vision Transformer 20 Jan 2022 · 0 repositories · arXiv:2201.08050
-
Transfer Learning Approaches for Building Cross-Language Dense Retrieval Models 20 Jan 2022 · 1 repository · arXiv:2201.08471Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
GAP-Gen: Guided Automatic Python Code Generation 19 Jan 2022 · 1 repository · arXiv:2201.08810
-
Near-Optimal Sparse Allreduce for Distributed Deep Learning 19 Jan 2022 · 1 repository · arXiv:2201.07598
-
Poseur: Direct Human Pose Regression with Transformers 19 Jan 2022 · 1 repository · arXiv:2201.07412
-
Q-ViT: Fully Differentiable Quantization for Vision Transformer 19 Jan 2022 · 1 repository · arXiv:2201.07703Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Swin-Pose: Swin Transformer Based Human Pose Estimation 19 Jan 2022 · 0 repositories · arXiv:2201.07384
-
TourBERT: A pretrained language model for the tourism industry 19 Jan 2022 · 0 repositories · arXiv:2201.07449
-
TransFuse: A Unified Transformer-based Image Fusion Framework using Self-supervised Learning 19 Jan 2022 · 0 repositories · arXiv:2201.07451
-
CoAuthor: Designing a Human-AI Collaborative Writing Dataset for Exploring Language Model Capabilities 18 Jan 2022 · 1 repository · arXiv:2201.06796
-
GTrans: Spatiotemporal Autoregressive Transformer with Graph Embeddings for Nowcasting Extreme Events 18 Jan 2022 · 0 repositories · arXiv:2201.06717
-
Hierarchical Neural Network Approaches for Long Document Classification 18 Jan 2022 · 0 repositories · arXiv:2201.06774
-
Motion Inbetweening via Deep Δ-Interpolator 18 Jan 2022 · 1 repository · arXiv:2201.06701