Methods › General › Attention Modules › Multi-Head Attention › Papers, page 174
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 174 of 249: papers 17,301 to 17,400 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Time Series Forecasting (TSF) Using Various Deep Learning Models 23 Apr 2022 · 0 repositories · arXiv:2204.11115
-
Attentions Help CNNs See Better: Attention-based Hybrid Image Quality Assessment Network 22 Apr 2022 · 3 repositories · arXiv:2204.10485
-
DFAM-DETR: Deformable feature based attention mechanism DETR on slender object detection 22 Apr 2022 · 0 repositories · arXiv:2204.10667
-
Diverse Instance Discovery: Vision-Transformer for Instance-Aware Multi-Label Image Recognition 22 Apr 2022 · 0 repositories · arXiv:2204.10731
-
End-to-end symbolic regression with transformers 22 Apr 2022 · 3 repositories · arXiv:2204.10532Syntology official (archive's flag): 1 ran · 4 ran (of which 1 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Fine-Tuning BERT Models to Classify Misinformation on Garlic and COVID-19 on Twitter 22 Apr 2022 · 0 repositories
-
Hierarchical Label-wise Attention Transformer Model for Explainable ICD Coding 22 Apr 2022 · 1 repository · arXiv:2204.10716
-
Hypergraph Transformer: Weakly-supervised Multi-hop Reasoning for Knowledge-based Visual Question Answering 22 Apr 2022 · 1 repository · arXiv:2204.10448
-
LibriS2S: A German-English Speech-to-Speech Translation Corpus 22 Apr 2022 · 1 repository · arXiv:2204.10593
-
Spatiality-guided Transformer for 3D Dense Captioning on Point Clouds 22 Apr 2022 · 1 repository · arXiv:2204.10688
-
Taygete at SemEval-2022 Task 4: RoBERTa based models for detecting Patronising and Condescending Language 22 Apr 2022 · 0 repositories · arXiv:2204.10519
-
Unified Pretraining Framework for Document Understanding 22 Apr 2022 · 0 repositories · arXiv:2204.10939
-
WaBERT: A Low-resource End-to-end Model for Spoken Language Understanding and Speech-to-BERT Alignment 22 Apr 2022 · 0 repositories · arXiv:2204.10461
-
BTranspose: Bottleneck Transformers for Human Pose Estimation with Self-Supervised Pre-Training 21 Apr 2022 · 0 repositories · arXiv:2204.10209
-
Measuring artificial intelligence: a systematic assessment and implications for governance 21 Apr 2022 · 0 repositories · arXiv:2204.10304
-
Is Neural Topic Modelling Better than Clustering? An Empirical Study on Clustering with Contextual Embeddings for Topics 21 Apr 2022 · 1 repository · arXiv:2204.09874Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Multi-Tier Platform for Cognizing Massive Electroencephalogram 21 Apr 2022 · 0 repositories · arXiv:2204.09840
-
SinTra: Learning an inspiration model from a single multi-track music segment 21 Apr 2022 · 1 repository · arXiv:2204.09917
-
Transformer-Guided Convolutional Neural Network for Cross-View Geolocalization 21 Apr 2022 · 0 repositories · arXiv:2204.09967
-
Detecting Unintended Memorization in Language-Model-Fused ASR 20 Apr 2022 · 0 repositories · arXiv:2204.09606
-
Human-Object Interaction Detection via Disentangled Transformer 20 Apr 2022 · 0 repositories · arXiv:2204.09290
-
Is BERT Robust to Label Noise? A Study on Learning with Noisy Labels in Text Classification 20 Apr 2022 · 1 repository · arXiv:2204.09371Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
NFormer: Robust Person Re-identification with Neighbor Transformer 20 Apr 2022 · 1 repository · arXiv:2204.09331Syntology official (archive's flag): 5 ran · 5 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 5 samples that ran constructed an object rather than computing a result (of 6 harvested samples)
-
Reinforced Structured State-Evolution for Vision-Language Navigation 20 Apr 2022 · 1 repository · arXiv:2204.09280
-
Towards Arabic Sentence Simplification via Classification and Generative Approaches 20 Apr 2022 · 0 repositories · arXiv:2204.09292
-
DiffMD: A Geometric Diffusion Model for Molecular Dynamics Simulations 19 Apr 2022 · 0 repositories · arXiv:2204.08672
-
ALBETO and DistilBETO: Lightweight Spanish Language Models 19 Apr 2022 · 2 repositories · arXiv:2204.09145
-
Blockwise Streaming Transformer for Spoken Language Understanding and Simultaneous Speech Translation 19 Apr 2022 · 0 repositories · arXiv:2204.08920
-
CodexDB: Generating Code for Processing SQL Queries using GPT-3 Codex 19 Apr 2022 · 0 repositories · arXiv:2204.08941
-
CTCNet: A CNN-Transformer Cooperation Network for Face Image Super-Resolution 19 Apr 2022 · 1 repository · arXiv:2204.08696
-
DecBERT: Enhancing the Language Understanding of BERT with Causal Attention Masks 19 Apr 2022 · 0 repositories · arXiv:2204.08688
-
Impact of Tokenization on Language Models: An Analysis for Turkish 19 Apr 2022 · 0 repositories · arXiv:2204.08832
-
LitMC-BERT: transformer-based multi-label classification of biomedical literature with an application on COVID-19 literature curation 19 Apr 2022 · 1 repository · arXiv:2204.08649
-
MANIQA: Multi-dimension Attention Network for No-Reference Image Quality Assessment 19 Apr 2022 · 2 repositories · arXiv:2204.08958Syntology official (archive's flag): 1 ran · 8 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 1 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
Mono vs Multilingual BERT for Hate Speech Detection and Text Classification: A Case Study in Marathi 19 Apr 2022 · 0 repositories · arXiv:2204.08669
-
Multi-View Spatial-Temporal Network for Continuous Sign Language Recognition 19 Apr 2022 · 0 repositories · arXiv:2204.08747
-
Multimodal Hate Speech Detection from Bengali Memes and Texts 19 Apr 2022 · 1 repository · arXiv:2204.10196
-
Not All Tokens Are Equal: Human-centric Visual Analysis via Token Clustering Transformer 19 Apr 2022 · 1 repository · arXiv:2204.08680Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Probing for the Usage of Grammatical Number 19 Apr 2022 · 0 repositories · arXiv:2204.08831
-
Self-Calibrated Efficient Transformer for Lightweight Super-Resolution 19 Apr 2022 · 1 repository · arXiv:2204.08913
-
Application of Transfer Learning and Ensemble Learning in Image-level Classification for Breast Histopathology 18 Apr 2022 · 0 repositories · arXiv:2204.08311
-
BSRT: Improving Burst Super-Resolution with Swin Transformer and Flow-Guided Deformable Alignment 18 Apr 2022 · 1 repository · arXiv:2204.08332Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 4 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples) · 2 pointer-only (licence)
-
Dynamic Position Encoding for Transformers 18 Apr 2022 · 0 repositories · arXiv:2204.08142
-
Ingredient Extraction from Text in the Recipe Domain 18 Apr 2022 · 1 repository · arXiv:2204.08137
-
L3Cube-HingCorpus and HingBERT: A Code Mixed Hindi-English Dataset and BERT Language Models 18 Apr 2022 · 1 repository · arXiv:2204.08398
-
MASSIVE: A 1M-Example Multilingual Natural Language Understanding Dataset with 51 Typologically-Diverse Languages 18 Apr 2022 · 6 repositories · arXiv:2204.08582Syntology official (archive's flag): 1 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Temporally Efficient Vision Transformer for Video Instance Segmentation 18 Apr 2022 · 3 repositories · arXiv:2204.08412Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
UTNLP at SemEval-2022 Task 6: A Comparative Analysis of Sarcasm Detection Using Generative-based and Mutation-based Data Augmentation 18 Apr 2022 · 2 repositories · arXiv:2204.08198
-
Visio-Linguistic Brain Encoding 18 Apr 2022 · 0 repositories · arXiv:2204.08261
-
Zero-shot Entity and Tweet Characterization with Designed Conditional Prompts and Contexts 18 Apr 2022 · 0 repositories · arXiv:2204.08405
-
An Extendable, Efficient and Effective Transformer-based Object Detector 17 Apr 2022 · 1 repository · arXiv:2204.07962Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Continual Hippocampus Segmentation with Transformers 17 Apr 2022 · 1 repository · arXiv:2204.08043
-
MST++: Multi-stage Spectral-wise Transformer for Efficient Spectral Reconstruction 17 Apr 2022 · 3 repositories · arXiv:2204.07908Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Knowledgeable Salient Span Mask for Enhancing Language Models as Knowledge Base 17 Apr 2022 · 0 repositories · arXiv:2204.07994
-
ParkPredict+: Multimodal Intent and Motion Prediction for Vehicles in Parking Lots with CNN and Transformer 17 Apr 2022 · 1 repository · arXiv:2204.10777
-
Pathologies of Pre-trained Language Models in Few-shot Fine-tuning 17 Apr 2022 · 0 repositories · arXiv:2204.08039
-
VDTR: Video Deblurring with Transformer 17 Apr 2022 · 1 repository · arXiv:2204.08023
-
A Hierarchical N-Gram Framework for Zero-Shot Link Prediction 16 Apr 2022 · 1 repository · arXiv:2204.10293
-
Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+ NLP Tasks 16 Apr 2022 · 10 repositories · arXiv:2204.07705Syntology community repositories only · 17 ran (of which 1 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 1 where Syntology's instrument failed) · 11 unverified (of 28 harvested samples) · 4 pointer-only (licence)
-
BLISS: Robust Sequence-to-Sequence Learning via Self-Supervised Input Representation 16 Apr 2022 · 0 repositories · arXiv:2204.07837
-
Privacy-Preserving Image Classification Using Isotropic Network 16 Apr 2022 · 0 repositories · arXiv:2204.07707
-
Probing Script Knowledge from Pre-Trained Models 16 Apr 2022 · 0 repositories · arXiv:2204.10176
-
Safe Self-Refinement for Transformer-based Domain Adaptation 16 Apr 2022 · 1 repository · arXiv:2204.07683Syntology official (archive's flag): 8 ran · 9 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 6 where Syntology's instrument failed) · 10 unverified (of 19 harvested samples) · 1 pointer-only (licence)
-
Searching Intrinsic Dimensions of Vision Transformers 16 Apr 2022 · 0 repositories · arXiv:2204.07722
-
SimpleBERT: A Pre-trained Model That Learns to Generate Simple Words 16 Apr 2022 · 0 repositories · arXiv:2204.07779
-
Towards Lightweight Transformer via Group-wise Transformation for Vision-and-Language Tasks 16 Apr 2022 · 1 repository · arXiv:2204.07780
-
WordAlchemy: A transformer-based Reverse Dictionary 16 Apr 2022 · 0 repositories · arXiv:2204.10181
-
ERGO: Event Relational Graph Transformer for Document-level Event Causality Identification 15 Apr 2022 · 0 repositories · arXiv:2204.07434
-
Image Captioning In the Transformer Age 15 Apr 2022 · 1 repository · arXiv:2204.07374
-
Improving Pre-trained Language Models with Syntactic Dependency Prediction Task for Chinese Semantic Error Recognition 15 Apr 2022 · 0 repositories · arXiv:2204.07464
-
mGPT: Few-Shot Learners Go Multilingual 15 Apr 2022 · 1 repository · arXiv:2204.07580
-
Mixture of Experts for Biomedical Question Answering 15 Apr 2022 · 0 repositories · arXiv:2204.07469
-
ML_LTU at SemEval-2022 Task 4: T5 Towards Identifying Patronizing and Condescending Language 15 Apr 2022 · 0 repositories · arXiv:2204.07432
-
MVSTER: Epipolar Transformer for Efficient Multi-View Stereo 15 Apr 2022 · 1 repository · arXiv:2204.07346Syntology official (archive's flag): 10 ran · 10 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 5 where Syntology's instrument failed) · 7 unverified (of 17 harvested samples)
-
On the Role of Pre-trained Language Models in Word Ordering: A Case Study with BART 15 Apr 2022 · 1 repository · arXiv:2204.07367
-
Polling Latent Opinions: A Method for Computational Sociolinguistics Using Transformer Language Models 15 Apr 2022 · 1 repository · arXiv:2204.07483
-
Pushing the Limits of Simple Pipelines for Few-Shot Learning: External Data and Fine-Tuning Make a Difference 15 Apr 2022 · 1 repository · arXiv:2204.07305Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 6 where Syntology's instrument failed) · 5 unverified (of 19 harvested samples) · 4 pointer-only (licence)
-
Resource-Aware Distributed Submodular Maximization: A Paradigm for Multi-Robot Decision-Making 15 Apr 2022 · 0 repositories · arXiv:2204.07520
-
ResT V2: Simpler, Faster and Stronger 15 Apr 2022 · 2 repositories · arXiv:2204.07366
-
Text Revision by On-the-Fly Representation Optimization 15 Apr 2022 · 1 repository · arXiv:2204.07359
-
Unconditional Image-Text Pair Generation with Multimodal Cross Quantizer 15 Apr 2022 · 1 repository · arXiv:2204.07537
-
3D Shuffle-Mixer: An Efficient Context-Aware Vision Learner of Transformer-MLP Paradigm for Dense Prediction in Medical Volume 14 Apr 2022 · 0 repositories · arXiv:2204.06779
-
Activation Regression for Continuous Domain Generalization with Applications to Crop Classification 14 Apr 2022 · 1 repository · arXiv:2204.07030
-
Analysing similarities between legal court documents using natural language processing approaches based on Transformers 14 Apr 2022 · 0 repositories · arXiv:2204.07182
-
CalBERT - Code-mixed Adaptive Language representations using BERT 14 Apr 2022 · 1 repository
-
Causal Transformer for Estimating Counterfactual Outcomes 14 Apr 2022 · 1 repository · arXiv:2204.07258
-
Challenges for Open-domain Targeted Sentiment Analysis 14 Apr 2022 · 0 repositories · arXiv:2204.06893
-
DeiT III: Revenge of the ViT 14 Apr 2022 · 12 repositories · arXiv:2204.07118
-
Does BERT really agree ? Fine-grained Analysis of Lexical Dependence on a Syntactic Task 14 Apr 2022 · 0 repositories · arXiv:2204.06889
-
Generative power of a protein language model trained on multiple sequence alignments 14 Apr 2022 · 1 repository · arXiv:2204.07110
-
GPT-NeoX-20B: An Open-Source Autoregressive Language Model 14 Apr 2022 · 11 repositories · arXiv:2204.06745Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Hierarchical Embedded Bayesian Additive Regression Trees 14 Apr 2022 · 0 repositories · arXiv:2204.07207
-
Latent Aspect Detection from Online Unsolicited Customer Reviews 14 Apr 2022 · 1 repository · arXiv:2204.06964
-
MiniViT: Compressing Vision Transformers with Weight Multiplexing 14 Apr 2022 · 2 repositories · arXiv:2204.07154Syntology official: harvested for another paper · 0 ran · 4 unverified (of 4 harvested samples)
-
Multi-label topic classification for COVID-19 literature with Bioformer 14 Apr 2022 · 0 repositories · arXiv:2204.06758
-
Residual Swin Transformer Channel Attention Network for Image Demosaicing 14 Apr 2022 · 0 repositories · arXiv:2204.07098
-
Rows from Many Sources: Enriching row completions from Wikidata with a pre-trained Language Model 14 Apr 2022 · 0 repositories · arXiv:2204.07014
-
Deep Relation Learning for Regression and Its Application to Brain Age Estimation 13 Apr 2022 · 0 repositories · arXiv:2204.06598
-
Fix Bugs with Transformer through a Neural-Symbolic Edit Grammar 13 Apr 2022 · 0 repositories · arXiv:2204.06643
-
Formal Language Recognition by Hard Attention Transformers: Perspectives from Circuit Complexity 13 Apr 2022 · 0 repositories · arXiv:2204.06618