Methods › General › Attention Mechanisms › Attention › Papers, page 234
Attention Is All You Need
Attention
Papers archive 2025-07-28
archive papers tagged: 31,583 · with a code link: 13,473 · where Syntology ran a sample: 3,998 (3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,998 of 31,583 tagged: 3,366 with a run with no instrument failure, 632 where every run was a failure of Syntology's instrument)
Page 234 of 316: papers 23,301 to 23,400 of 31,583, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
CoSIm: Commonsense Reasoning for Counterfactual Scene Imagination 8 Jul 2022 · 1 repository · arXiv:2207.03961
-
Cross-Attention Transformer for Video Interpolation 8 Jul 2022 · 1 repository · arXiv:2207.04132
-
Deep Visual-Linguistic Fusion Network Considering Cross-Modal Inconsistency for Rumor Detection 8 Jul 2022 · 3 repositories
-
Getting BART to Ride the Idiomatic Train: Learning to Represent Idiomatic Expressions 8 Jul 2022 · 1 repository · arXiv:2207.03679
-
Hidden Schema Networks 8 Jul 2022 · 0 repositories · arXiv:2207.03777
-
RePFormer: Refinement Pyramid Transformer for Robust Facial Landmark Detection 8 Jul 2022 · 0 repositories · arXiv:2207.03917
-
VidConv: A modernized 2D ConvNet for Efficient Video Recognition 8 Jul 2022 · 1 repository · arXiv:2207.03782
-
A Large Scale Search Dataset for Unbiased Learning to Rank 7 Jul 2022 · 1 repository · arXiv:2207.03051
-
Active Learning and Multi-label Classification for Ellipsis and Coreference Detection in Conversational Question-Answering 7 Jul 2022 · 0 repositories · arXiv:2207.03145
-
AsNER -- Annotated Dataset and Baseline for Assamese Named Entity recognition 7 Jul 2022 · 0 repositories · arXiv:2207.03422
-
Back to the Basics: Revisiting Out-of-Distribution Detection Baselines 7 Jul 2022 · 2 repositories · arXiv:2207.03061Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Deformer: Towards Displacement Field Learning for Unsupervised Medical Image Registration 7 Jul 2022 · 1 repository · arXiv:2207.03180
-
Deep learning based Hand gesture recognition system and design of a Human-Machine Interface 7 Jul 2022 · 0 repositories · arXiv:2207.03112
-
Dual-Stream Transformer for Generic Event Boundary Captioning 7 Jul 2022 · 1 repository · arXiv:2207.03038
-
Mirror Complementary Transformer Network for RGB-thermal Salient Object Detection 7 Jul 2022 · 1 repository · arXiv:2207.03558
-
More ConvNets in the 2020s: Scaling up Kernels Beyond 51x51 using Sparsity 7 Jul 2022 · 1 repository · arXiv:2207.03620
-
Neural Language Models are not Born Equal to Fit Brain Data, but Training Helps 7 Jul 2022 · 0 repositories · arXiv:2207.03380
-
Sensitivity Analysis on Transferred Neural Architectures of BERT and GPT-2 for Financial Sentiment Analysis 7 Jul 2022 · 0 repositories · arXiv:2207.03037
-
Ask Me What You Need: Product Retrieval using Knowledge from GPT-3 6 Jul 2022 · 0 repositories · arXiv:2207.02516
-
Aspect-Based Sentiment Analysis using Local Context Focus Mechanism with DeBERTa 6 Jul 2022 · 0 repositories · arXiv:2207.02424
-
Astroconformer: Inferring Surface Gravity of Stars from Stellar Light Curves with Transformer 6 Jul 2022 · 0 repositories · arXiv:2207.02787
-
Branchformer: Parallel MLP-Attention Architectures to Capture Local and Global Context for Speech Recognition and Understanding 6 Jul 2022 · 3 repositories · arXiv:2207.02971
-
Cross-receptive Focused Inference Network for Lightweight Image Super-Resolution 6 Jul 2022 · 1 repository · arXiv:2207.02796
-
Delving into Sequential Patches for Deepfake Detection 6 Jul 2022 · 0 repositories · arXiv:2207.02803
-
Don't Pay Attention to the Noise: Learning Self-supervised Representations of Light Curves with a Denoising Time Series Transformer 6 Jul 2022 · 1 repository · arXiv:2207.02777Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 2 honoured, 0 violated, 10 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
FAST-VQA: Efficient End-to-end Video Quality Assessment with Fragment Sampling 6 Jul 2022 · 4 repositories · arXiv:2207.02595Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Learning to Diversify for Product Question Generation 6 Jul 2022 · 0 repositories · arXiv:2207.02534
-
MaiT: Leverage Attention Masks for More Efficient Image Transformers 6 Jul 2022 · 0 repositories · arXiv:2207.03006
-
Pre-training Transformers for Molecular Property Prediction Using Reaction Prediction 6 Jul 2022 · 0 repositories · arXiv:2207.02724
-
Pure Transformers are Powerful Graph Learners 6 Jul 2022 · 2 repositories · arXiv:2207.02505Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 2 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
SimLM: Pre-training with Representation Bottleneck for Dense Passage Retrieval 6 Jul 2022 · 1 repository · arXiv:2207.02578
-
The Role of Complex NLP in Transformers for Text Ranking? 6 Jul 2022 · 0 repositories · arXiv:2207.02522
-
Transformers are Adaptable Task Planners 6 Jul 2022 · 0 repositories · arXiv:2207.02442
-
Betti numbers of attention graphs is all you really need 5 Jul 2022 · 1 repository · arXiv:2207.01903
-
CNN-based Local Vision Transformer for COVID-19 Diagnosis 5 Jul 2022 · 0 repositories · arXiv:2207.02027
-
CoBEVT: Cooperative Bird's Eye View Semantic Segmentation with Sparse Transformers 5 Jul 2022 · 2 repositories · arXiv:2207.02202Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
CodeRL: Mastering Code Generation through Pretrained Models and Deep Reinforcement Learning 5 Jul 2022 · 2 repositories · arXiv:2207.01780Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Deep Learning Reveals Patterns of Diverse and Changing Sentiments Towards COVID-19 Vaccines Based on 11 Million Tweets 5 Jul 2022 · 0 repositories · arXiv:2207.10641
-
Detecting and Recovering Sequential DeepFake Manipulation 5 Jul 2022 · 1 repository · arXiv:2207.02204Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
FishFormer: Annulus Slicing-based Transformer for Fisheye Rectification with Efficacy Domain Exploration 5 Jul 2022 · 0 repositories · arXiv:2207.01925
-
Improving Semantic Segmentation in Transformers using Hierarchical Inter-Level Attention 5 Jul 2022 · 0 repositories · arXiv:2207.02126
-
Machine Learning Model Sizes and the Parameter Gap 5 Jul 2022 · 0 repositories · arXiv:2207.02852
-
TabPFN: A Transformer That Solves Small Tabular Classification Problems in a Second 5 Jul 2022 · 7 repositories · arXiv:2207.01848Syntology 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Multimodal Frame-Scoring Transformer for Video Summarization 5 Jul 2022 · 0 repositories · arXiv:2207.01814
-
Softmax-free Linear Transformers 5 Jul 2022 · 1 repository · arXiv:2207.03341
-
Swin Deformable Attention U-Net Transformer (SDAUT) for Explainable Fast MRI 5 Jul 2022 · 1 repository · arXiv:2207.02390
-
Transformer based Models for Unsupervised Anomaly Segmentation in Brain MR Images 5 Jul 2022 · 0 repositories · arXiv:2207.02059
-
Ultra-Low-Bitrate Speech Coding with Pretrained Transformers 5 Jul 2022 · 0 repositories · arXiv:2207.02262
-
An adaptive music generation architecture for games based on the deep learning Transformer mode 4 Jul 2022 · 0 repositories · arXiv:2207.01698
-
BERT, can HE predict contrastive focus? Predicting and controlling prominence in neural TTS using a language model 4 Jul 2022 · 0 repositories · arXiv:2207.01718
-
Interaction Transformer for Human Reaction Generation 4 Jul 2022 · 1 repository · arXiv:2207.01685
-
Large-scale Robustness Analysis of Video Action Recognition Models 4 Jul 2022 · 1 repository · arXiv:2207.01398
-
One Model is Not Enough: Ensembles for Isolated Sign Language Recognition 4 Jul 2022 · 0 repositories
-
Towards Real-World Video Denosing: A Practical Video Denosing Dataset and Network 4 Jul 2022 · 0 repositories · arXiv:2207.01356
-
TANet: Transformer-based Asymmetric Network for RGB-D Salient Object Detection 4 Jul 2022 · 1 repository · arXiv:2207.01172
-
Understanding Performance of Long-Document Ranking Models through Comprehensive Evaluation and Leaderboarding 4 Jul 2022 · 3 repositories · arXiv:2207.01262
-
Using contextual sentence analysis models to recognize ESG concepts 4 Jul 2022 · 0 repositories · arXiv:2207.01402
-
Divert More Attention to Vision-Language Tracking 3 Jul 2022 · 1 repository · arXiv:2207.01076
-
Generating Repetitions with Appropriate Repeated Words 3 Jul 2022 · 1 repository · arXiv:2207.00929
-
Unified Object Detector for Different Modalities based on Vision Transformers 3 Jul 2022 · 1 repository · arXiv:2207.01071
-
GUIM -- General User and Item Embedding with Mixture of Representation in E-commerce 2 Jul 2022 · 0 repositories · arXiv:2207.00750
-
Improving Transformer-based Conversational ASR by Inter-Sentential Attention Mechanism 2 Jul 2022 · 0 repositories · arXiv:2207.00883
-
A Polyphone BERT for Polyphone Disambiguation in Mandarin Chinese 1 Jul 2022 · 0 repositories · arXiv:2207.12089
-
A Temporal Fusion Transformer for Long-term Explainable Prediction of Emergency Department Overcrowding 1 Jul 2022 · 0 repositories · arXiv:2207.00610
-
DALG: Deep Attentive Local and Global Modeling for Image Retrieval 1 Jul 2022 · 0 repositories · arXiv:2207.00287
-
Dissecting Self-Supervised Learning Methods for Surgical Computer Vision 1 Jul 2022 · 1 repository · arXiv:2207.00449
-
Improving Low-Resource Speech Recognition with Pretrained Speech Models: Continued Pretraining vs. Semi-Supervised Training 1 Jul 2022 · 0 repositories · arXiv:2207.00659
-
Is neural language acquisition similar to natural? A chronological probing study 1 Jul 2022 · 1 repository · arXiv:2207.00560
-
Masked Autoencoder for Self-Supervised Pre-training on Lidar Point Clouds 1 Jul 2022 · 1 repository · arXiv:2207.00531Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Polarized Color Image Denoising using Pocoformer 1 Jul 2022 · 0 repositories · arXiv:2207.00215
-
Time-aware Dynamic Graph Embedding for Asynchronous Structural Evolution 1 Jul 2022 · 0 repositories · arXiv:2207.00594
-
TopicFM: Robust and Interpretable Topic-Assisted Feature Matching 1 Jul 2022 · 1 repository · arXiv:2207.00328
-
AnoShift: A Distribution Shift Benchmark for Unsupervised Anomaly Detection 30 Jun 2022 · 1 repository · arXiv:2206.15476Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
Classical and learned MR to pseudo-CT mappings for accurate transcranial ultrasound simulation 30 Jun 2022 · 0 repositories · arXiv:2206.15441
-
Compressing Pre-trained Transformers via Low-Bit NxM Sparsity for Natural Language Understanding 30 Jun 2022 · 0 repositories · arXiv:2206.15014
-
CTrGAN: Cycle Transformers GAN for Gait Transfer 30 Jun 2022 · 0 repositories · arXiv:2206.15248
-
Deep Reinforcement Learning with Swin Transformers 30 Jun 2022 · 1 repository · arXiv:2206.15269
-
DeepSpeed Inference: Enabling Efficient Inference of Transformer Models at Unprecedented Scale 30 Jun 2022 · 2 repositories · arXiv:2207.00032
-
FL-Tuning: Layer Tuning for Feed-Forward Network in Transformer 30 Jun 2022 · 1 repository · arXiv:2206.15312
-
GaitForeMer: Self-Supervised Pre-Training of Transformers via Human Motion Forecasting for Few-Shot Gait Impairment Severity Estimation 30 Jun 2022 · 1 repository · arXiv:2207.00106
-
No Reason for No Supervision: Improved Generalization in Supervised Models 30 Jun 2022 · 1 repository · arXiv:2206.15369Syntology 7 ran (of which 6 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
ListBERT: Learning to Rank E-commerce products with Listwise BERT 30 Jun 2022 · 0 repositories · arXiv:2206.15198
-
PolarFormer: Multi-camera 3D Object Detection with Polar Transformer 30 Jun 2022 · 1 repository · arXiv:2206.15398Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
PVT-COV19D: Pyramid Vision Transformer for COVID-19 Diagnosis 30 Jun 2022 · 0 repositories · arXiv:2206.15069
-
Rethinking Surgical Captioning: End-to-End Window-Based MLP Transformer Using Patches 30 Jun 2022 · 1 repository · arXiv:2207.00113
-
TENET: Transformer Encoding Network for Effective Temporal Flow on Motion Prediction 30 Jun 2022 · 0 repositories · arXiv:2207.00170
-
The Topological BERT: Transforming Attention into Topology for Natural Language Processing 30 Jun 2022 · 0 repositories · arXiv:2206.15195
-
Two-Stage Classifier for COVID-19 Misinformation Detection Using BERT: a Study on Indonesian Tweets 30 Jun 2022 · 2 repositories · arXiv:2206.15359
-
BATFormer: Towards Boundary-Aware Lightweight Transformer for Efficient Medical Image Segmentation 29 Jun 2022 · 2 repositories · arXiv:2206.14409
-
Chinese Word Sense Embedding with SememeWSD and Synonym Set 29 Jun 2022 · 2 repositories · arXiv:2206.14388
-
Deformable Graph Transformer 29 Jun 2022 · 0 repositories · arXiv:2206.14337
-
Not Cheating on the Turing Test: Towards Grounded Language Learning in Artificial Intelligence 29 Jun 2022 · 0 repositories · arXiv:2206.14672
-
LViT: Language meets Vision Transformer in Medical Image Segmentation 29 Jun 2022 · 1 repository · arXiv:2206.14718Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Multi-Channel Vision Transformer for Epileptic Seizure Prediction 29 Jun 2022 · 0 repositories
-
On the Prediction Network Architecture in RNN-T for ASR 29 Jun 2022 · 0 repositories · arXiv:2206.14618
-
SALO: An Efficient Spatial Accelerator Enabling Hybrid Sparse Attention Mechanisms for Long Sequences 29 Jun 2022 · 0 repositories · arXiv:2206.14550
-
Simple and Effective Multi-sentence TTS with Expressive and Coherent Prosody 29 Jun 2022 · 0 repositories · arXiv:2206.14643
-
Two-Stage COVID19 Classification Using BERT Features 29 Jun 2022 · 0 repositories · arXiv:2206.14861
-
Cross-Forgery Analysis of Vision Transformers and CNNs for Deepfake Image Detection 28 Jun 2022 · 2 repositories · arXiv:2206.13829
-
Exploring linguistic feature and model combination for speech recognition based automatic AD detection 28 Jun 2022 · 0 repositories · arXiv:2206.13758