Methods › General › Attention Modules › Multi-Head Attention › Papers, page 166
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 166 of 249: papers 16,501 to 16,600 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Scaling Laws vs Model Architectures: How does Inductive Bias Influence Scaling? 21 Jul 2022 · 0 repositories · arXiv:2207.10551
-
SeedFormer: Patch Seeds based Point Cloud Completion with Upsample Transformer 21 Jul 2022 · 1 repository · arXiv:2207.10315Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Sequence Models for Drone vs Bird Classification 21 Jul 2022 · 0 repositories · arXiv:2207.10409
-
The Birth of Bias: A case study on the evolution of gender bias in an English language model 21 Jul 2022 · 1 repository · arXiv:2207.10245
-
Towards Efficient Adversarial Training on Vision Transformers 21 Jul 2022 · 0 repositories · arXiv:2207.10498
-
Weakly Supervised Object Localization via Transformer with Implicit Spatial Calibration 21 Jul 2022 · 2 repositories · arXiv:2207.10447Syntology official (archive's flag): 7 ran · 7 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)
-
AiATrack: Attention in Attention for Transformer Visual Tracking 20 Jul 2022 · 1 repository · arXiv:2207.09603Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; the one sample that ran constructed an object rather than computing a result (of 3 harvested samples)
-
HTNet: Anchor-free Temporal Action Localization with Hierarchical Transformers 20 Jul 2022 · 0 repositories · arXiv:2207.09662
-
Locality Guidance for Improving Vision Transformers on Tiny Datasets 20 Jul 2022 · 1 repository · arXiv:2207.10026Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
MeshMAE: Masked Autoencoders for 3D Mesh Data Analysis 20 Jul 2022 · 0 repositories · arXiv:2207.10228
-
Unsupervised Industrial Anomaly Detection via Pattern Generative and Contrastive Networks 20 Jul 2022 · 0 repositories · arXiv:2207.09792
-
ViGAT: Bottom-up event recognition and explanation in video using factorized graph attention network 20 Jul 2022 · 1 repository · arXiv:2207.09927
-
Enhancing Collaborative Filtering Recommender with Prompt-Based Sentiment Analysis 19 Jul 2022 · 1 repository · arXiv:2207.12883
-
GAFX: A General Audio Feature eXtractor 19 Jul 2022 · 0 repositories · arXiv:2207.09145
-
Investigation of deep learning models on identification of minimum signal length for precise classification of conveyor rubber belt loads 19 Jul 2022 · 1 repository
-
PiC: A Phrase-in-Context Dataset for Phrase Understanding and Semantic Search 19 Jul 2022 · 1 repository · arXiv:2207.09068
-
Pre-trained language models with domain knowledge for biomedical extractive summarization 19 Jul 2022 · 1 repository
-
Revealing Secrets From Pre-trained Models 19 Jul 2022 · 0 repositories · arXiv:2207.09539
-
Target-Driven Structured Transformer Planner for Vision-Language Navigation 19 Jul 2022 · 1 repository · arXiv:2207.11201
-
TTVFI: Learning Trajectory-Aware Transformer for Video Frame Interpolation 19 Jul 2022 · 0 repositories · arXiv:2207.09048
-
Vision Transformers: From Semantic Segmentation to Dense Prediction 19 Jul 2022 · 3 repositories · arXiv:2207.09339
-
AlexU-AIC at Arabic Hate Speech 2022: Contrast to Classify 18 Jul 2022 · 0 repositories · arXiv:2207.08557
-
Conditional DETR V2: Efficient Detection Transformer with Box Queries 18 Jul 2022 · 0 repositories · arXiv:2207.08914
-
Dense Cross-Query-and-Support Attention Weighted Mask Aggregation for Few-Shot Segmentation 18 Jul 2022 · 1 repository · arXiv:2207.08549Syntology official (archive's flag): 4 ran · 4 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Focal-WNet: An Architecture Unifying Convolution and Attention for Depth Estimation 18 Jul 2022 · 1 repository
-
HiFormer: Hierarchical Multi-scale Representations Using Transformers for Medical Image Segmentation 18 Jul 2022 · 1 repository · arXiv:2207.08518
-
Multi-manifold Attention for Vision Transformers 18 Jul 2022 · 0 repositories · arXiv:2207.08569
-
Selection Bias Induced Spurious Correlations in Large Language Models 18 Jul 2022 · 1 repository · arXiv:2207.08982
-
TokenMix: Rethinking Image Mixing for Data Augmentation in Vision Transformers 18 Jul 2022 · 1 repository · arXiv:2207.08409Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Word Play for Playing Othello (Reverses) 18 Jul 2022 · 0 repositories · arXiv:2207.08766
-
A Multibias-mitigated and Sentiment Knowledge Enriched Transformer for Debiasing in Multimodal Conversational Emotion Recognition 17 Jul 2022 · 0 repositories · arXiv:2207.08104
-
Aspect-specific Context Modeling for Aspect-based Sentiment Analysis 17 Jul 2022 · 1 repository · arXiv:2207.08099
-
Can large language models reason about medical questions? 17 Jul 2022 · 1 repository · arXiv:2207.08143Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples)
-
Defect Transformer: An Efficient Hybrid Transformer Architecture for Surface Defect Detection 17 Jul 2022 · 0 repositories · arXiv:2207.08319
-
Effectiveness of French Language Models on Abstractive Dialogue Summarization Task 17 Jul 2022 · 0 repositories · arXiv:2207.08305
-
ELECTRA is a Zero-Shot Learner, Too 17 Jul 2022 · 1 repository · arXiv:2207.08141Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
FashionViL: Fashion-Focused Vision-and-Language Representation Learning 17 Jul 2022 · 1 repository · arXiv:2207.08150
-
MDM: Multiple Dynamic Masks for Visual Explanation of Neural Networks 17 Jul 2022 · 0 repositories · arXiv:2207.08046
-
Representation Learning of Image Schema 17 Jul 2022 · 0 repositories · arXiv:2207.08256
-
Robust Action Governor for Uncertain Piecewise Affine Systems with Non-convex Constraints and Safe Reinforcement Learning 17 Jul 2022 · 0 repositories · arXiv:2207.08240
-
A Context-Sensitive Word Embedding Approach for The Detection of Troll Tweets 17 Jul 2022 · 0 repositories · arXiv:2207.08230
-
CharFormer: A Glyph Fusion based Attentive Framework for High-precision Character Image Denoising 16 Jul 2022 · 1 repository · arXiv:2207.07798
-
Explainable vision transformer enabled convolutional neural network for plant disease identification: PlantXViT 16 Jul 2022 · 0 repositories · arXiv:2207.07919
-
Mitigating Data Redundancy to Revitalize Transformer-based Long-Term Time Series Forecasting System 16 Jul 2022 · 2 repositories · arXiv:2207.07827
-
Generative Adversarial Networks Based on Transformer Encoder and Convolution Block for Hyperspectral Image Classification 16 Jul 2022 · 0 repositories
-
JPerceiver: Joint Perception Network for Depth, Pose and Layout Estimation in Driving Scenes 16 Jul 2022 · 1 repository · arXiv:2207.07895Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Multimodal Dialog Systems with Dual Knowledge-enhanced Generative Pretrained Language Model 16 Jul 2022 · 0 repositories · arXiv:2207.07934
-
SSMTL++: Revisiting Self-Supervised Multi-Task Learning for Video Anomaly Detection 16 Jul 2022 · 0 repositories · arXiv:2207.08003
-
Structural Prior Guided Generative Adversarial Transformers for Low-Light Image Enhancement 16 Jul 2022 · 0 repositories · arXiv:2207.07828
-
A Systematic Review and Replicability Study of BERT4Rec for Sequential Recommendation 15 Jul 2022 · 1 repository · arXiv:2207.07483Syntology official (archive's flag): 2 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
CKD-TransBTS: Clinical Knowledge-Driven Hybrid Transformer with Modality-Correlated Cross-Attention for Brain Tumor Segmentation 15 Jul 2022 · 0 repositories · arXiv:2207.07370
-
Lightweight Vision Transformer with Cross Feature Attention 15 Jul 2022 · 0 repositories · arXiv:2207.07268
-
Mobile Keystroke Biometrics Using Transformers 15 Jul 2022 · 1 repository · arXiv:2207.07596
-
POET: Training Neural Networks on Tiny Devices with Integrated Rematerialization and Paging 15 Jul 2022 · 1 repository · arXiv:2207.07697
-
Position Prediction as an Effective Pretraining Strategy 15 Jul 2022 · 1 repository · arXiv:2207.07611
-
Z-Index at CheckThat! Lab 2022: Check-Worthiness Identification on Tweet Text 15 Jul 2022 · 0 repositories · arXiv:2207.07308
-
Combing for Credentials: Active Pattern Extraction from Smart Reply 14 Jul 2022 · 0 repositories · arXiv:2207.10802
-
Bootstrapped Masked Autoencoders for Vision BERT Pretraining 14 Jul 2022 · 1 repository · arXiv:2207.07116Syntology official (archive's flag): 6 ran · 6 ran (of which 6 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 6 samples that ran constructed an object rather than computing a result (of 7 harvested samples) · 7 pointer-only (licence)
-
Convolutional Bypasses Are Better Vision Transformer Adapters 14 Jul 2022 · 1 repository · arXiv:2207.07039Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Current Trends in Deep Learning for Earth Observation: An Open-source Benchmark Arena for Image Classification 14 Jul 2022 · 2 repositories · arXiv:2207.07189
-
Deepfake Video Detection with Spatiotemporal Dropout Transformer 14 Jul 2022 · 0 repositories · arXiv:2207.06612
-
Forming Trees with Treeformers 14 Jul 2022 · 0 repositories · arXiv:2207.06960
-
iColoriT: Towards Propagating Local Hint to the Right Region in Interactive Colorization by Leveraging Vision Transformer 14 Jul 2022 · 1 repository · arXiv:2207.06831
-
Language Modelling with Pixels 14 Jul 2022 · 1 repository · arXiv:2207.06991Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Multilinguals at SemEval-2022 Task 11: Complex NER in Semantically Ambiguous Settings for Low Resource Languages 14 Jul 2022 · 1 repository · arXiv:2207.06882
-
Multitrack Music Transformer 14 Jul 2022 · 2 repositories · arXiv:2207.06983Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples)
-
Recurrent Memory Transformer 14 Jul 2022 · 3 repositories · arXiv:2207.06881Syntology official (archive's flag): 3 ran · 6 ran (of which 3 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 1 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 7 unverified (of 13 harvested samples) · 2 pointer-only (licence)
-
Scene Text Recognition with Permuted Autoregressive Sequence Models 14 Jul 2022 · 2 repositories · arXiv:2207.06966Syntology official (archive's flag): 5 ran · 5 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified; every one of the 5 samples that ran constructed an object rather than computing a result (of 8 harvested samples)
-
Transformer-based Context Condensation for Boosting Feature Pyramids in Object Detection 14 Jul 2022 · 0 repositories · arXiv:2207.06603
-
A Transfer Learning Based Model for Text Readability Assessment in German 13 Jul 2022 · 0 repositories · arXiv:2207.06265
-
Diverse Dance Synthesis via Keyframes with Transformer Controllers 13 Jul 2022 · 1 repository · arXiv:2207.05906
-
DocPrompting: Generating Code by Retrieving the Docs 13 Jul 2022 · 2 repositories · arXiv:2207.05987Syntology official (archive's flag): 4 ran · 4 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 9 harvested samples)
-
DynaST: Dynamic Sparse Transformer for Exemplar-Guided Image Generation 13 Jul 2022 · 1 repository · arXiv:2207.06124Syntology official (archive's flag): 5 ran · 5 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples)
-
Entry-Flipped Transformer for Inference and Prediction of Participant Behavior 13 Jul 2022 · 0 repositories · arXiv:2207.06235
-
Exploiting Word Semantics to Enrich Character Representations of Chinese Pre-trained Models 13 Jul 2022 · 1 repository · arXiv:2207.05928
-
Fuse It More Deeply! A Variational Transformer with Layer-Wise Latent Variable Inference for Text Generation 13 Jul 2022 · 1 repository · arXiv:2207.06130Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 3 pointer-only (licence)
-
Masked Autoencoders that Listen 13 Jul 2022 · 4 repositories · arXiv:2207.06405Syntology official (archive's flag): 1 ran · 17 ran (of which 12 constructed an object rather than computing a result; 16 with no instrument failure: 1 honoured, 1 violated, 14 with no contract checked; 1 where Syntology's instrument failed) · 15 unverified (of 32 harvested samples) · 10 pointer-only (licence)
-
N-Grammer: Augmenting Transformers with latent n-grams 13 Jul 2022 · 2 repositories · arXiv:2207.06366Syntology official (archive's flag): 5 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Pyramid Transformer for Traffic Sign Detection 13 Jul 2022 · 0 repositories · arXiv:2207.06067
-
Re2G: Retrieve, Rerank, Generate 13 Jul 2022 · 1 repository · arXiv:2207.06300Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
RTN: Reinforced Transformer Network for Coronary CT Angiography Vessel-level Image Quality Assessment 13 Jul 2022 · 0 repositories · arXiv:2207.06177
-
Unsupervised Visual Representation Learning by Synchronous Momentum Grouping 13 Jul 2022 · 1 repository · arXiv:2207.06167Syntology 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
Visual Context-driven Audio Feature Enhancement for Robust End-to-End Audio-Visual Speech Recognition 13 Jul 2022 · 1 repository · arXiv:2207.06020
-
A new hope for network model generalization 12 Jul 2022 · 1 repository · arXiv:2207.05843
-
Cross-Architecture Knowledge Distillation 12 Jul 2022 · 0 repositories · arXiv:2207.05273
-
Earthformer: Exploring Space-Time Transformers for Earth System Forecasting 12 Jul 2022 · 2 repositories · arXiv:2207.05833Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples)
-
eX-ViT: A Novel eXplainable Vision Transformer for Weakly Supervised Semantic Segmentation 12 Jul 2022 · 0 repositories · arXiv:2207.05358
-
How Do Multilingual Encoders Learn Cross-lingual Representation? 12 Jul 2022 · 0 repositories · arXiv:2207.05737
-
Image and Model Transformation with Secret Key for Vision Transformer 12 Jul 2022 · 0 repositories · arXiv:2207.05366
-
MSP-Former: Multi-Scale Projection Transformer for Single Image Desnowing 12 Jul 2022 · 0 repositories · arXiv:2207.05621
-
Multi-Behavior Hypergraph-Enhanced Transformer for Sequential Recommendation 12 Jul 2022 · 1 repository · arXiv:2207.05584Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Next-ViT: Next Generation Vision Transformer for Efficient Deployment in Realistic Industrial Scenarios 12 Jul 2022 · 5 repositories · arXiv:2207.05501Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
OSLAT: Open Set Label Attention Transformer for Medical Entity Retrieval and Span Extraction 12 Jul 2022 · 1 repository · arXiv:2207.05817
-
Dateformer: Time-modeling Transformer for Longer-term Series Forecasting 12 Jul 2022 · 1 repository · arXiv:2207.05397
-
Towards Hard-Positive Query Mining for DETR-based Human-Object Interaction Detection 12 Jul 2022 · 1 repository · arXiv:2207.05293Syntology official (archive's flag): 5 ran · 5 ran (of which 2 constructed an object rather than computing a result; 4 with no instrument failure: 2 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Trusted Multi-Scale Classification Framework for Whole Slide Image 12 Jul 2022 · 0 repositories · arXiv:2207.05290
-
Using Paraphrases to Study Properties of Contextual Embeddings 12 Jul 2022 · 0 repositories · arXiv:2207.05553
-
Video Graph Transformer for Video Question Answering 12 Jul 2022 · 1 repository · arXiv:2207.05342Syntology official (archive's flag): 9 ran · 9 ran (of which 5 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 14 harvested samples)
-
Dual Vision Transformer 11 Jul 2022 · 1 repository · arXiv:2207.04976
-
Learning Large-scale Universal User Representation with Sparse Mixture of Experts 11 Jul 2022 · 0 repositories · arXiv:2207.04648