Methods › General › Attention Modules › Multi-Head Attention › Papers, page 187
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 187 of 249: papers 18,601 to 18,700 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Text classification problems via BERT embedding method and graph convolutional neural network 30 Nov 2021 · 0 repositories · arXiv:2111.15379
-
Text Mining Drug/Chemical-Protein Interactions using an Ensemble of BERT and T5 Based Models 30 Nov 2021 · 0 repositories · arXiv:2111.15617
-
Building extraction with vision transformer 29 Nov 2021 · 0 repositories · arXiv:2111.15637
-
Customer Sentiment Analysis using Weak Supervision for Customer-Agent Chat 29 Nov 2021 · 0 repositories · arXiv:2111.14282
-
DAFormer: Improving Network Architectures and Training Strategies for Domain-Adaptive Semantic Segmentation 29 Nov 2021 · 3 repositories · arXiv:2111.14887Syntology official: no sample here; runs from other or unrecorded repositories · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
End-to-End Referring Video Object Segmentation with Multimodal Transformers 29 Nov 2021 · 2 repositories · arXiv:2111.14821Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 6 pointer-only (licence)
-
Mixed Precision Low-bit Quantization of Neural Network Language Models for Speech Recognition 29 Nov 2021 · 0 repositories · arXiv:2112.11438
-
Mixed Precision of Quantization of Transformer Language Models for Speech Recognition 29 Nov 2021 · 0 repositories · arXiv:2112.11540
-
On the rate of convergence of a classifier based on a Transformer encoder 29 Nov 2021 · 0 repositories · arXiv:2111.14574
-
Point-BERT: Pre-training 3D Point Cloud Transformers with Masked Point Modeling 29 Nov 2021 · 3 repositories · arXiv:2111.14819Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Recurrent Vision Transformer for Solving Visual Reasoning Problems 29 Nov 2021 · 0 repositories · arXiv:2111.14576
-
Searching the Search Space of Vision Transformer 29 Nov 2021 · 2 repositories · arXiv:2111.14725
-
Sparse DETR: Efficient End-to-End Object Detection with Learnable Sparsity 29 Nov 2021 · 1 repository · arXiv:2111.14330Syntology official (archive's flag): 3 ran · 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 7 harvested samples)
-
Speech Tasks Relevant to Sleepiness Determined with Deep Transfer Learning 29 Nov 2021 · 0 repositories · arXiv:2111.14684
-
TransMVSNet: Global Context-aware Multi-view Stereo Network with Transformers 29 Nov 2021 · 1 repository · arXiv:2111.14600
-
Context Matters in Semantically Controlled Language Generation for Task-oriented Dialogue Systems 28 Nov 2021 · 0 repositories · arXiv:2111.14119
-
FastTrees: Parallel Latent Tree-Induction for Faster Sequence Encoding 28 Nov 2021 · 1 repository · arXiv:2111.14031
-
Multi-domain Integrative Swin Transformer network for Sparse-View Tomographic Reconstruction 28 Nov 2021 · 0 repositories · arXiv:2111.14831
-
ORCHARD: A Benchmark For Measuring Systematic Generalization of Multi-Hierarchical Reasoning 28 Nov 2021 · 1 repository · arXiv:2111.14034
-
Abusive and Threatening Language Detection in Urdu using Boosting based and BERT based models: A Comparative Approach 27 Nov 2021 · 1 repository · arXiv:2111.14830
-
Exploring Transformer Based Models to Identify Hate Speech and Offensive Content in English and Indo-Aryan Languages 27 Nov 2021 · 0 repositories · arXiv:2111.13974
-
FQ-ViT: Post-Training Quantization for Fully Quantized Vision Transformer 27 Nov 2021 · 1 repository · arXiv:2111.13824Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
GEOSCAN: Global Earth Observation using Swarm of Coordinated Autonomous Nanosats 27 Nov 2021 · 0 repositories · arXiv:2111.15627
-
Learning A 3D-CNN and Transformer Prior for Hyperspectral Image Super-Resolution 27 Nov 2021 · 0 repositories · arXiv:2111.13923
-
Tapping BERT for Preposition Sense Disambiguation 27 Nov 2021 · 0 repositories · arXiv:2111.13972
-
A Robust Volumetric Transformer for Accurate 3D Tumor Segmentation 26 Nov 2021 · 1 repository · arXiv:2111.13300Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Exploiting full Resolution Feature Context for Liver Tumor and Vessel Segmentation via Integrate Framework: Application to Liver Tumor and Vessel 3D Reconstruction under embedded microprocessor 26 Nov 2021 · 1 repository · arXiv:2111.13299
-
GMFlow: Learning Optical Flow via Global Matching 26 Nov 2021 · 4 repositories · arXiv:2111.13680Syntology official (archive's flag): 11 ran · 14 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 19 harvested samples) · 7 pointer-only (licence)
-
Predicting Document Coverage for Relation Extraction 26 Nov 2021 · 0 repositories · arXiv:2111.13611
-
SWAT: Spatial Structure Within and Among Tokens 26 Nov 2021 · 1 repository · arXiv:2111.13677
-
Domain Prompt Learning for Efficiently Adapting CLIP to Unseen Domains 25 Nov 2021 · 1 repository · arXiv:2111.12853Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
Attend to Who You Are: Supervising Self-Attention for Keypoint Detection and Instance-Aware Association 25 Nov 2021 · 1 repository · arXiv:2111.12892
-
BoxeR: Box-Attention for 2D and 3D Transformers 25 Nov 2021 · 1 repository · arXiv:2111.13087
-
Does constituency analysis enhance domain-specific pre-trained BERT models for relation extraction? 25 Nov 2021 · 0 repositories · arXiv:2112.02955
-
Evaluating the Robustness of Retrieval Pipelines with Query Variation Generators 25 Nov 2021 · 1 repository · arXiv:2111.13057
-
Exploiting Both Domain-specific and Invariant Knowledge via a Win-win Transformer for Unsupervised Domain Adaptation 25 Nov 2021 · 0 repositories · arXiv:2111.12941
-
Global Interaction Modelling in Vision Transformer via Super Tokens 25 Nov 2021 · 0 repositories · arXiv:2111.13156
-
New Approaches to Long Document Summarization: Fourier Transform Based Attention in a Transformer Model 25 Nov 2021 · 0 repositories · arXiv:2111.15473
-
Probabilistic Impact Score Generation using Ktrain-BERT to Identify Hate Words from Twitter Discussions 25 Nov 2021 · 0 repositories · arXiv:2111.12939
-
Recommending Multiple Positive Citations for Manuscript via Content-Dependent Modeling and Multi-Positive Triplet 25 Nov 2021 · 0 repositories · arXiv:2111.12899
-
Scene Representation Transformer: Geometry-Free Novel View Synthesis Through Set-Latent Scene Representations 25 Nov 2021 · 1 repository · arXiv:2111.13152Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Transformer-based Korean Pretrained Language Models: A Survey on Three Years of Progress 25 Nov 2021 · 0 repositories · arXiv:2112.03014
-
TunBERT: Pretrained Contextualized Text Representation for Tunisian Dialect 25 Nov 2021 · 0 repositories · arXiv:2111.13138
-
Attention-based Dual-stream Vision Transformer for Radar Gait Recognition 24 Nov 2021 · 0 repositories · arXiv:2111.12290
-
MHFormer: Multi-Hypothesis Transformer for 3D Human Pose Estimation 24 Nov 2021 · 1 repository · arXiv:2111.12707Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
MorphMLP: An Efficient MLP-Like Backbone for Spatial-Temporal Representation Learning 24 Nov 2021 · 2 repositories · arXiv:2111.12527Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
PeCo: Perceptual Codebook for BERT Pre-training of Vision Transformers 24 Nov 2021 · 1 repository · arXiv:2111.12710
-
Sparse is Enough in Scaling Transformers 24 Nov 2021 · 0 repositories · arXiv:2111.12763
-
Unleashing Transformers: Parallel Token Prediction with Discrete Absorbing Diffusion for Fast High-Resolution Image Generation from Vector-Quantized Codes 24 Nov 2021 · 3 repositories · arXiv:2111.12701
-
Utilizing Resource-Rich Language Datasets for End-to-End Scene Text Recognition in Resource-Poor Languages 24 Nov 2021 · 0 repositories · arXiv:2111.12276
-
VIOLET : End-to-End Video-Language Transformers with Masked Visual-token Modeling 24 Nov 2021 · 1 repository · arXiv:2111.12681Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
DABS: A Domain-Agnostic Benchmark for Self-Supervised Learning 23 Nov 2021 · 1 repository · arXiv:2111.12062Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Deep Point Cloud Reconstruction 23 Nov 2021 · 0 repositories · arXiv:2111.11704
-
Multi-Person 3D Motion Prediction with Multi-Range Transformers 23 Nov 2021 · 1 repository · arXiv:2111.12073
-
Pruning Self-attentions into Convolutional Layers in Single Path 23 Nov 2021 · 3 repositories · arXiv:2111.11802Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
SimpleTRON: Simple Transformer with O(N) Complexity 23 Nov 2021 · 0 repositories · arXiv:2111.15588
-
S-SimCSE: Sampled Sub-networks for Contrastive Learning of Sentence Embedding 23 Nov 2021 · 0 repositories · arXiv:2111.11750
-
Self-Supervised Pre-Training for Transformer-Based Person Re-Identification 23 Nov 2021 · 3 repositories · arXiv:2111.12084Syntology community repositories only · 7 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
U-shape Transformer for Underwater Image Enhancement 23 Nov 2021 · 2 repositories · arXiv:2111.11843Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Variational Learning for Unsupervised Knowledge Grounded Dialogs 23 Nov 2021 · 1 repository · arXiv:2112.00653
-
Benchmarking Detection Transfer Learning with Vision Transformers 22 Nov 2021 · 2 repositories · arXiv:2111.11429
-
DBIA: Data-free Backdoor Injection Attack against Transformer Networks 22 Nov 2021 · 1 repository · arXiv:2111.11870
-
ExT5: Towards Extreme Multi-Task Scaling for Transfer Learning 22 Nov 2021 · 4 repositories · arXiv:2111.10952Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples)
-
Can depth-adaptive BERT perform better on binary classification tasks 22 Nov 2021 · 0 repositories · arXiv:2111.10951
-
L-Verse: Bidirectional Generation Between Image and Text 22 Nov 2021 · 1 repository · arXiv:2111.11133
-
MetaFormer Is Actually What You Need for Vision 22 Nov 2021 · 18 repositories · arXiv:2111.11418Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Patch Vestiges in the Adversarial Examples Against Vision Transformer Can Be Leveraged for Adversarial Detection 22 Nov 2021 · 0 repositories
-
Lightweight Transformer Backbone for Medical Object Detection 22 Nov 2021 · 0 repositories · arXiv:2111.11546
-
Semi-Supervised Vision Transformers 22 Nov 2021 · 1 repository · arXiv:2111.11067
-
CpT: Convolutional Point Transformer for 3D Point Cloud Processing 21 Nov 2021 · 0 repositories · arXiv:2111.10866
-
DuDoTrans: Dual-Domain Transformer Provides More Attention for Sinogram Restoration in Sparse-View CT Reconstruction 21 Nov 2021 · 0 repositories · arXiv:2111.10790
-
Efficient Softmax Approximation for Deep Neural Networks with Attention Mechanism 21 Nov 2021 · 0 repositories · arXiv:2111.10770
-
Are Vision Transformers Robust to Patch Perturbations? 20 Nov 2021 · 0 repositories · arXiv:2111.10659
-
Data Processing Matters: SRPH-Konvergen AI's Machine Translation System for WMT'21 20 Nov 2021 · 0 repositories · arXiv:2111.10513
-
Discrete Representations Strengthen Vision Transformer Robustness 20 Nov 2021 · 1 repository · arXiv:2111.10493
-
Advancing High-Resolution Video-Language Representation with Large-Scale Video Transcriptions 19 Nov 2021 · 1 repository · arXiv:2111.10337
-
Does BERT look at sentiment lexicon? 19 Nov 2021 · 0 repositories · arXiv:2111.10100
-
DyFormer: A Scalable Dynamic Graph Transformer with Provable Benefits on Generalization Ability 19 Nov 2021 · 0 repositories · arXiv:2111.10447
-
Generalized Decision Transformer for Offline Hindsight Information Matching 19 Nov 2021 · 1 repository · arXiv:2111.10364Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Grounded Situation Recognition with Transformers 19 Nov 2021 · 1 repository · arXiv:2111.10135Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Lexicon-based Methods vs. BERT for Text Sentiment Analysis 19 Nov 2021 · 0 repositories · arXiv:2111.10097
-
Rethinking Query, Key, and Value Embedding in Vision Transformer under Tiny Model Constraints 19 Nov 2021 · 0 repositories · arXiv:2111.10017
-
TransMorph: Transformer for unsupervised medical image registration 19 Nov 2021 · 2 repositories · arXiv:2111.10480Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
PatchCensor: Patch Robustness Certification for Transformers via Exhaustive Testing 19 Nov 2021 · 0 repositories · arXiv:2111.10481
-
ClipCap: CLIP Prefix for Image Captioning 18 Nov 2021 · 4 repositories · arXiv:2111.09734Syntology official (archive's flag): 5 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing 18 Nov 2021 · 3 repositories · arXiv:2111.09543Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Dynamic-TinyBERT: Boost TinyBERT's Inference Efficiency by Dynamic Sequence Length 18 Nov 2021 · 0 repositories · arXiv:2111.09645
-
How Emotionally Stable is ALBERT? Testing Robustness with Stochastic Weight Averaging on a Sentiment Analysis Task 18 Nov 2021 · 1 repository · arXiv:2111.09612
-
How much do language models copy from their training data? Evaluating linguistic novelty in text generation using RAVEN 18 Nov 2021 · 0 repositories · arXiv:2111.09509
-
LAnoBERT: System Log Anomaly Detection based on BERT Masked Language Model 18 Nov 2021 · 0 repositories · arXiv:2111.09564
-
Multimodal Emotion Recognition on RAVDESS Dataset Using Transfer Learning 18 Nov 2021 · 0 repositories
-
Reference-based Magnetic Resonance Image Reconstruction Using Texture Transformer 18 Nov 2021 · 0 repositories · arXiv:2111.09492
-
Restormer: Efficient Transformer for High-Resolution Image Restoration 18 Nov 2021 · 13 repositories · arXiv:2111.09881Syntology official (archive's flag): 2 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
RoBERTuito: a pre-trained language model for social media text in Spanish 18 Nov 2021 · 1 repository · arXiv:2111.09453Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
Swin Transformer V2: Scaling Up Capacity and Resolution 18 Nov 2021 · 23 repositories · arXiv:2111.09883Syntology community repositories only · 17 ran (of which 0 constructed an object rather than computing a result; 17 with no instrument failure: 0 honoured, 3 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 13 unverified (of 30 harvested samples) · 4 pointer-only (licence)
-
The Power of Selecting Key Blocks with Local Pre-ranking for Long Document Information Retrieval 18 Nov 2021 · 1 repository · arXiv:2111.09852
-
You Only Sample (Almost) Once: Linear Cost Self-Attention Via Bernoulli Sampling 18 Nov 2021 · 1 repository · arXiv:2111.09714Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Guiding Generative Language Models for Data Augmentation in Few-Shot Text Classification 17 Nov 2021 · 0 repositories · arXiv:2111.09064
-
A Comparative Study on Transfer Learning and Distance Metrics in Semantic Clustering over the COVID-19 Tweets 16 Nov 2021 · 0 repositories · arXiv:2111.08658
-
A Deep Generative XAI Framework for Natural Language Inference Explanations Generation 16 Nov 2021 · 0 repositories