Methods › General › Attention Mechanisms › Attention › Papers, page 253
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 253 of 316: papers 25,201 to 25,300 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
A deep language model to predict metabolic network equilibria 7 Dec 2021 · 0 repositories · arXiv:2112.03588
-
A Transferable Approach for Partitioning Machine Learning Models on Multi-Chip-Modules 7 Dec 2021 · 0 repositories · arXiv:2112.04041
-
Attention-Based Model and Deep Reinforcement Learning for Distribution of Event Processing Tasks 7 Dec 2021 · 1 repository · arXiv:2112.03835
-
Bootstrapping ViTs: Towards Liberating Vision Transformers from Pre-training 7 Dec 2021 · 1 repository · arXiv:2112.03552
-
Emulating Spatio-Temporal Realizations of Three-Dimensional Isotropic Turbulence via Deep Sequence Learning Models 7 Dec 2021 · 1 repository · arXiv:2112.03469
-
raceBERT -- A Transformer-based Model for Predicting Race and Ethnicity from Names 7 Dec 2021 · 1 repository · arXiv:2112.03807
-
Regularity Learning via Explicit Distribution Modeling for Skeletal Video Anomaly Detection 7 Dec 2021 · 1 repository · arXiv:2112.03649
-
Relating transformers to models and neural representations of the hippocampal formation 7 Dec 2021 · 0 repositories · arXiv:2112.04035
-
SSAT: A Symmetric Semantic-Aware Transformer Network for Makeup Transfer and Removal 7 Dec 2021 · 2 repositories · arXiv:2112.03631
-
GETAM: Gradient-weighted Element-wise Transformer Attention Map for Weakly-supervised Semantic segmentation 6 Dec 2021 · 1 repository · arXiv:2112.02841
-
Offline Pre-trained Multi-Agent Decision Transformer: One Big Sequence Model Tackles All SMAC Tasks 6 Dec 2021 · 1 repository · arXiv:2112.02845
-
One-shot Talking Face Generation from Single-speaker Audio-Visual Correlation Learning 6 Dec 2021 · 0 repositories · arXiv:2112.02749
-
PTTR: Relational 3D Point Cloud Object Tracking with Transformer 6 Dec 2021 · 1 repository · arXiv:2112.02857
-
Scaling Up Influence Functions 6 Dec 2021 · 2 repositories · arXiv:2112.03052
-
Skeletal Graph Self-Attention: Embedding a Skeleton Inductive Bias into Sign Language Production 6 Dec 2021 · 0 repositories · arXiv:2112.05277
-
Spatio-Temporal meets Wavelet: Disentangled Traffic Flow Forecasting via Efficient Spectral Graph Attention Network 6 Dec 2021 · 0 repositories · arXiv:2112.02740
-
Team Hitachi @ AutoMin 2021: Reference-free Automatic Minuting Pipeline with Argument Structure Construction over Topic-based Summarization 6 Dec 2021 · 0 repositories · arXiv:2112.02741
-
BERTMap: A BERT-based Ontology Alignment System 5 Dec 2021 · 1 repository · arXiv:2112.02682
-
Causal Distillation for Language Models 5 Dec 2021 · 1 repository · arXiv:2112.02505Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
DIBERT: Dependency Injected Bidirectional Encoder Representations from Transformers 5 Dec 2021 · 1 repository
-
Dynamic Token Normalization Improves Vision Transformers 5 Dec 2021 · 1 repository · arXiv:2112.02624Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Gaudí: Conversational Interactions with Deep Representations to Generate Image Collections 5 Dec 2021 · 0 repositories · arXiv:2112.04404
-
Learning Tracking Representations via Dual-Branch Fully Transformer Networks 5 Dec 2021 · 1 repository · arXiv:2112.02571
-
PolyphonicFormer: Unified Query Learning for Depth-aware Video Panoptic Segmentation 5 Dec 2021 · 1 repository · arXiv:2112.02582
-
Pose-guided Feature Disentangling for Occluded Person Re-identification Based on Transformer 5 Dec 2021 · 1 repository · arXiv:2112.02466
-
VarCLR: Variable Semantic Representation Pre-training via Contrastive Learning 5 Dec 2021 · 1 repository · arXiv:2112.02650Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
3rd Place: A Global and Local Dual Retrieval Solution to Facebook AI Image Similarity Challenge 4 Dec 2021 · 1 repository · arXiv:2112.02373
-
A Multi-Strategy based Pre-Training Method for Cold-Start Recommendation 4 Dec 2021 · 0 repositories · arXiv:2112.02275
-
Bridging Pre-trained Models and Downstream Tasks for Source Code Understanding 4 Dec 2021 · 1 repository · arXiv:2112.02268Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples)
-
LAVT: Language-Aware Vision Transformer for Referring Image Segmentation 4 Dec 2021 · 1 repository · arXiv:2112.02244
-
Representation Learning for Conversational Data using Discourse Mutual Information Maximization 4 Dec 2021 · 0 repositories · arXiv:2112.05787
-
U2-Former: A Nested U-shaped Transformer for Image Restoration 4 Dec 2021 · 0 repositories · arXiv:2112.02279
-
Unraveling Social Perceptions & Behaviors towards Migrants on Twitter 4 Dec 2021 · 0 repositories · arXiv:2112.06642
-
A Novel Deep Parallel Time-series Relation Network for Fault Diagnosis 3 Dec 2021 · 0 repositories · arXiv:2112.03405
-
Augmenting Customer Support with an NLP-based Receptionist 3 Dec 2021 · 0 repositories · arXiv:2112.01959
-
CTIN: Robust Contextual Transformer Network for Inertial Navigation 3 Dec 2021 · 1 repository · arXiv:2112.02143
-
Efficient Two-Stage Detection of Human-Object Interactions with a Novel Unary-Pairwise Transformer 3 Dec 2021 · 1 repository · arXiv:2112.01838Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Given Users Recommendations Based on Reviews on Yelp 3 Dec 2021 · 1 repository · arXiv:2112.01762
-
Make A Long Image Short: Adaptive Token Length for Vision Transformers 3 Dec 2021 · 0 repositories · arXiv:2112.01686
-
NN-LUT: Neural Approximation of Non-Linear Operations for Efficient Transformer Inference 3 Dec 2021 · 0 repositories · arXiv:2112.02191
-
Siamese BERT-based Model for Web Search Relevance Ranking Evaluated on a New Czech Dataset 3 Dec 2021 · 1 repository · arXiv:2112.01810
-
Single-Shot Black-Box Adversarial Attacks Against Malware Detectors: A Causal Language Model Approach 3 Dec 2021 · 0 repositories · arXiv:2112.01724
-
TransZero: Attribute-guided Transformer for Zero-Shot Learning 3 Dec 2021 · 1 repository · arXiv:2112.01683
-
BEVT: BERT Pretraining of Video Transformers 2 Dec 2021 · 1 repository · arXiv:2112.01529
-
Masked-attention Mask Transformer for Universal Image Segmentation 2 Dec 2021 · 7 repositories · arXiv:2112.01527Syntology community repositories only · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples)
-
MutualFormer: Multi-Modality Representation Learning via Cross-Diffusion Attention 2 Dec 2021 · 1 repository · arXiv:2112.01177
-
PLSUM: Generating PT-BR Wikipedia by Summarizing Multiple Websites 2 Dec 2021 · 1 repository · arXiv:2112.01591
-
ScaleVLAD: Improving Multimodal Sentiment Analysis via Multi-Scale Fusion of Locally Descriptors 2 Dec 2021 · 0 repositories · arXiv:2112.01368
-
Self-supervised Video Transformer 2 Dec 2021 · 1 repository · arXiv:2112.01514Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
SwinTrack: A Simple and Strong Baseline for Transformer Tracking 2 Dec 2021 · 1 repository · arXiv:2112.00995Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
TBN-ViT: Temporal Bilateral Network with Vision Transformer for Video Scene Parsing 2 Dec 2021 · 0 repositories · arXiv:2112.01033
-
PTCT: Patches with 3D-Temporal Convolutional Transformer Network for Precipitation Nowcasting 2 Dec 2021 · 1 repository · arXiv:2112.01085
-
Uni-Perceiver: Pre-training Unified Architecture for Generic Perception for Zero-shot and Few-shot Tasks 2 Dec 2021 · 1 repository · arXiv:2112.01522
-
Unsupervised Law Article Mining based on Deep Pre-Trained Language Representation Models with Application to the Italian Civil Code 2 Dec 2021 · 0 repositories · arXiv:2112.03033
-
Visual-Semantic Transformer for Scene Text Recognition 2 Dec 2021 · 0 repositories · arXiv:2112.00948
-
Co-evolution Transformer for Protein Contact Prediction 1 Dec 2021 · 1 repository
-
Combining Global and Local Attention with Positional Encoding for Video Summarization 1 Dec 2021 · 1 repository
-
Container: Context Aggregation Networks 1 Dec 2021 · 2 repositories
-
Controlling Conditional Language Models without Catastrophic Forgetting 1 Dec 2021 · 2 repositories · arXiv:2112.00791Syntology official (archive's flag): 3 ran · 7 ran (of which 1 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 4 pointer-only (licence)
-
Cross-view Geo-localization with Layer-to-Layer Transformer 1 Dec 2021 · 0 repositories
-
Detecting Extratropical Cyclones of the Northern Hemisphere with Single Shot Detector 1 Dec 2021 · 0 repositories · arXiv:2112.01283
-
Do Transformers Really Perform Badly for Graph Representation? 1 Dec 2021 · 0 repositories
-
Domain-oriented Language Pre-training with Adaptive Hybrid Masking and Optimal Transport Alignment 1 Dec 2021 · 0 repositories · arXiv:2112.03024
-
DRONE: Data-aware Low-rank Compression for Large NLP Models 1 Dec 2021 · 0 repositories
-
Federated Split Task-Agnostic Vision Transformer for COVID-19 CXR Diagnosis 1 Dec 2021 · 0 repositories
-
Focal Attention for Long-Range Interactions in Vision Transformers 1 Dec 2021 · 1 repository
-
Gauge Equivariant Transformer 1 Dec 2021 · 0 repositories
-
HRFormer: High-Resolution Vision Transformer for Dense Predict 1 Dec 2021 · 2 repositories
-
Integrating Tree Path in Transformer for Code Representation 1 Dec 2021 · 1 repository
-
Multi-View Stereo with Transformer 1 Dec 2021 · 0 repositories · arXiv:2112.00336
-
NER-BERT: A Pre-trained Model for Low-Resource Entity Tagging 1 Dec 2021 · 0 repositories · arXiv:2112.00405
-
Raw Nav-merge Seismic Data to Subsurface Properties with MLP based Multi-Modal Information Unscrambler 1 Dec 2021 · 0 repositories
-
Score Transformer: Generating Musical Score from Note-level Representation 1 Dec 2021 · 1 repository · arXiv:2112.00355
-
Searching for Efficient Transformers for Language Modeling 1 Dec 2021 · 0 repositories
-
Shapeshifter: a Parameter-efficient Transformer using Factorized Reshaped Matrices 1 Dec 2021 · 1 repository
-
Speech-T: Transducer for Text to Speech and Beyond 1 Dec 2021 · 0 repositories
-
Systematic Generalization with Edge Transformers 1 Dec 2021 · 1 repository · arXiv:2112.00578
-
TEDGE-Caching: Transformer-based Edge Caching Towards 6G Networks 1 Dec 2021 · 0 repositories · arXiv:2112.00633
-
Think Big, Teach Small: Do Language Models Distil Occam’s Razor? 1 Dec 2021 · 1 repository
-
TriBERT: Human-centric Audio-visual Representation Learning 1 Dec 2021 · 1 repository
-
Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer 1 Dec 2021 · 1 repository
-
UniDoc: Unified Pretraining Framework for Document Understanding 1 Dec 2021 · 0 repositories
-
Wiki to Automotive: Understanding the Distribution Shift and its impact on Named Entity Recognition 1 Dec 2021 · 0 repositories · arXiv:2112.00283
-
A Comparative Study of Transformers on Word Sense Disambiguation 30 Nov 2021 · 0 repositories · arXiv:2111.15417
-
A Unified Pruning Framework for Vision Transformers 30 Nov 2021 · 1 repository · arXiv:2111.15127
-
Adaptive Token Sampling For Efficient Vision Transformers 30 Nov 2021 · 1 repository · arXiv:2111.15667Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Boosting Discriminative Visual Representation Learning with Scenario-Agnostic Mixup 30 Nov 2021 · 1 repository · arXiv:2111.15454
-
Chemical Identification and Indexing in PubMed Articles via BERT and Text-to-Text Approaches 30 Nov 2021 · 0 repositories · arXiv:2111.15622
-
Generating Rich Product Descriptions for Conversational E-commerce Systems 30 Nov 2021 · 0 repositories · arXiv:2111.15298
-
HEAT: Holistic Edge Attention Transformer for Structured Reconstruction 30 Nov 2021 · 1 repository · arXiv:2111.15143
-
KARL-Trans-NER: Knowledge Aware Representation Learning for Named Entity Recognition using Transformers 30 Nov 2021 · 0 repositories · arXiv:2111.15436
-
NLP Techniques for Water Quality Analysis in Social Media Content 30 Nov 2021 · 0 repositories · arXiv:2112.11441
-
Pixelated Butterfly: Simple and Efficient Sparse training for Neural Network Models 30 Nov 2021 · 1 repository · arXiv:2112.00029Syntology official: harvested, nothing ran · 4 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 4 samples that ran constructed an object rather than computing a result (of 6 harvested samples) · 5 pointer-only (licence)
-
Pyramid Adversarial Training Improves ViT Performance 30 Nov 2021 · 1 repository · arXiv:2111.15121
-
Robust Partial-to-Partial Point Cloud Registration in a Full Range 30 Nov 2021 · 1 repository · arXiv:2111.15606
-
Sentiment Analysis and Effect of COVID-19 Pandemic using College SubReddit Data 30 Nov 2021 · 1 repository · arXiv:2112.04351
-
Shunted Self-Attention via Multi-Scale Token Aggregation 30 Nov 2021 · 1 repository · arXiv:2111.15193
-
SpaceEdit: Learning a Unified Editing Space for Open-Domain Image Editing 30 Nov 2021 · 0 repositories · arXiv:2112.00180
-
Text classification problems via BERT embedding method and graph convolutional neural network 30 Nov 2021 · 0 repositories · arXiv:2111.15379
-
Text Mining Drug/Chemical-Protein Interactions using an Ensemble of BERT and T5 Based Models 30 Nov 2021 · 0 repositories · arXiv:2111.15617