Methods › General › Attention Modules › Multi-Head Attention › Papers, page 165
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 165 of 249: papers 16,401 to 16,500 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Combined CNN Transformer Encoder for Enhanced Fine-grained Human Action Recognition 3 Aug 2022 · 0 repositories · arXiv:2208.01897
-
Efficient Fine-Tuning of Compressed Language Models with Learners 3 Aug 2022 · 0 repositories · arXiv:2208.02070
-
Multi-Feature Vision Transformer via Self-Supervised Representation Learning for Improvement of COVID-19 Diagnosis 3 Aug 2022 · 1 repository · arXiv:2208.01843
-
SSformer: A Lightweight Transformer for Semantic Segmentation 3 Aug 2022 · 1 repository · arXiv:2208.02034
-
A Comparative Study on COVID-19 Fake News Detection Using Different Transformer Based Models 2 Aug 2022 · 0 repositories · arXiv:2208.01355
-
A Novel Transformer Network with Shifted Window Cross-Attention for Spatiotemporal Weather Forecasting 2 Aug 2022 · 0 repositories · arXiv:2208.01252
-
Active entailment encoding for explanation tree construction using parsimonious generation of hard negatives 2 Aug 2022 · 0 repositories · arXiv:2208.01376
-
Automatic Classification of Bug Reports Based on Multiple Text Information and Reports' Intention 2 Aug 2022 · 0 repositories · arXiv:2208.01274
-
Debiasing Gender Bias in Information Retrieval Models 2 Aug 2022 · 0 repositories · arXiv:2208.01755
-
Making the Best of Both Worlds: A Domain-Oriented Transformer for Unsupervised Domain Adaptation 2 Aug 2022 · 1 repository · arXiv:2208.01195
-
Multi-Module G2P Converter for Persian Focusing on Relations between Words 2 Aug 2022 · 0 repositories · arXiv:2208.01371
-
Two-Stream Transformer Architecture for Long Video Understanding 2 Aug 2022 · 0 repositories · arXiv:2208.01753
-
BATMAN: Bilateral Attention Transformer in Motion-Appearance Neighboring Space for Video Object Segmentation 1 Aug 2022 · 0 repositories · arXiv:2208.01159
-
giMLPs: Gate with Inhibition Mechanism in MLPs 1 Aug 2022 · 1 repository · arXiv:2208.00929
-
Local Perception-Aware Transformer for Aerial Tracking 1 Aug 2022 · 1 repository · arXiv:2208.00662
-
Physics-inform attention temporal convolutional network for EEG-based motor imagery classification 1 Aug 2022 · 2 repositories
-
Pose Uncertainty Aware Movement Synchrony Estimation via Spatial-Temporal Graph Transformer 1 Aug 2022 · 0 repositories · arXiv:2208.01161
-
SiamixFormer: a fully-transformer Siamese network with temporal Fusion for accurate building detection and change detection in bi-temporal remote sensing images 1 Aug 2022 · 0 repositories · arXiv:2208.00657
-
Interacting with next-phrase suggestions: How suggestion systems aid and influence the cognitive processes of writing 1 Aug 2022 · 0 repositories · arXiv:2208.00636
-
TransDeepLab: Convolution-Free Transformer-based DeepLab v3+ for Medical Image Segmentation 1 Aug 2022 · 1 repository · arXiv:2208.00713
-
What Can Transformers Learn In-Context? A Case Study of Simple Function Classes 1 Aug 2022 · 2 repositories · arXiv:2208.01066
-
Aggretriever: A Simple Approach to Aggregate Textual Representations for Robust Dense Passage Retrieval 31 Jul 2022 · 1 repository · arXiv:2208.00511
-
Neural Knowledge Bank for Pretrained Transformers 31 Jul 2022 · 0 repositories · arXiv:2208.00399
-
One for All: One-stage Referring Expression Comprehension with Dynamic Reasoning 31 Jul 2022 · 0 repositories · arXiv:2208.00361
-
STrajNet: Multi-modal Hierarchical Transformer for Occupancy Flow Field Prediction in Autonomous Driving 31 Jul 2022 · 1 repository · arXiv:2208.00394
-
Toward Understanding WordArt: Corner-Guided Transformer for Scene Text Recognition 31 Jul 2022 · 1 repository · arXiv:2208.00438Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
A Survey on Masked Autoencoder for Self-supervised Learning in Vision and Beyond 30 Jul 2022 · 0 repositories · arXiv:2208.00173
-
Doubly Deformable Aggregation of Covariance Matrices for Few-shot Segmentation 30 Jul 2022 · 1 repository · arXiv:2208.00306Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples) · 2 pointer-only (licence)
-
ALADIN: Distilling Fine-grained Alignment Scores for Efficient Image-Text Matching and Retrieval 29 Jul 2022 · 1 repository · arXiv:2207.14757
-
Code Comment Inconsistency Detection with BERT and Longformer 29 Jul 2022 · 1 repository · arXiv:2207.14444
-
Curriculum Learning for Data-Efficient Vision-Language Alignment 29 Jul 2022 · 0 repositories · arXiv:2207.14525
-
FCSN: Global Context Aware Segmentation by Learning the Fourier Coefficients of Objects in Medical Images 29 Jul 2022 · 0 repositories · arXiv:2207.14477
-
Forensic License Plate Recognition with Compression-Informed Transformers 29 Jul 2022 · 0 repositories · arXiv:2207.14686
-
GTrans: Grouping and Fusing Transformer Layers for Neural Machine Translation 29 Jul 2022 · 1 repository · arXiv:2207.14467Syntology official (archive's flag): 4 ran · 6 ran (of which 2 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 3 pointer-only (licence)
-
SERCNN: Stacked Embedding Recurrent Convolutional Neural Network in Detecting Depression on Twitter 29 Jul 2022 · 0 repositories · arXiv:2207.14535
-
Towards Unconstrained Audio Splicing Detection and Localization with Neural Networks 29 Jul 2022 · 0 repositories · arXiv:2207.14682
-
CrAM: A Compression-Aware Minimizer 28 Jul 2022 · 1 repository · arXiv:2207.14200Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples)
-
Depth Field Networks for Generalizable Multi-view Scene Representation 28 Jul 2022 · 1 repository · arXiv:2207.14287
-
DnSwin: Toward Real-World Denoising via Continuous Wavelet Sliding-Transformer 28 Jul 2022 · 1 repository · arXiv:2207.13861
-
LAD: Language Models as Data for Zero-Shot Dialog 28 Jul 2022 · 0 repositories · arXiv:2207.14393
-
Large Language Models and the Reverse Turing Test 28 Jul 2022 · 0 repositories · arXiv:2207.14382
-
Neural Architecture Search on Efficient Transformers and Beyond 28 Jul 2022 · 0 repositories · arXiv:2207.13955
-
SDBERT: SparseDistilBERT, a faster and smaller BERT model 28 Jul 2022 · 0 repositories · arXiv:2208.10246
-
Self-Supervised Hypergraph Transformer for Recommender Systems 28 Jul 2022 · 1 repository · arXiv:2207.14338
-
Semantic-Aligned Matching for Enhanced DETR Convergence and Multi-Scale Feature Fusion 28 Jul 2022 · 1 repository · arXiv:2207.14172
-
Sequence to sequence pretraining for a less-resourced Slovenian language 28 Jul 2022 · 1 repository · arXiv:2207.13988
-
Subtype-Former: a deep learning approach for cancer subtype discovery with multi-omics data 28 Jul 2022 · 0 repositories · arXiv:2207.14639
-
A Variational AutoEncoder for Transformers with Nonparametric Variational Information Bottleneck 27 Jul 2022 · 0 repositories · arXiv:2207.13529
-
Are Neighbors Enough? Multi-Head Neural n-gram can be Alternative to Self-attention 27 Jul 2022 · 0 repositories · arXiv:2207.13354
-
AvatarPoser: Articulated Full-Body Pose Tracking from Sparse Motion Sensing 27 Jul 2022 · 1 repository · arXiv:2207.13784Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 3 pointer-only (licence)
-
Cross-Attention of Disentangled Modalities for 3D Human Mesh Recovery with Transformers 27 Jul 2022 · 1 repository · arXiv:2207.13820Syntology official (archive's flag): 12 ran · 12 ran (of which 4 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 4 where Syntology's instrument failed) · 4 unverified (of 16 harvested samples)
-
Deep Clustering with Features from Self-Supervised Pretraining 27 Jul 2022 · 0 repositories · arXiv:2207.13364
-
Is Attention All That NeRF Needs? 27 Jul 2022 · 1 repository · arXiv:2207.13298Syntology 5 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 5 samples that ran constructed an object rather than computing a result (of 7 harvested samples)
-
SoundChoice: Grapheme-to-Phoneme Models with Semantic Disambiguation 27 Jul 2022 · 1 repository · arXiv:2207.13703
-
TransNorm: Transformer Provides a Strong Spatial Normalization Mechanism for a Deep Segmentation Model 27 Jul 2022 · 1 repository · arXiv:2207.13415
-
Bodily Behaviors in Social Interaction: Novel Annotations and State-of-the-Art Evaluation 26 Jul 2022 · 0 repositories · arXiv:2207.12817
-
Bundle MCR: Towards Conversational Bundle Recommendation 26 Jul 2022 · 1 repository · arXiv:2207.12628
-
Cross-Modal Causal Relational Reasoning for Event-Level Visual Question Answering 26 Jul 2022 · 2 repositories · arXiv:2207.12647Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
DETRs with Hybrid Matching 26 Jul 2022 · 8 repositories · arXiv:2207.13080Syntology official (archive's flag): 15 ran · 16 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 2 honoured, 2 violated, 5 with no contract checked; 7 where Syntology's instrument failed) · 2 unverified (of 18 harvested samples) · 4 pointer-only (licence)
-
Graph Neural Network and Spatiotemporal Transformer Attention for 3D Video Object Detection from Point Clouds 26 Jul 2022 · 0 repositories · arXiv:2207.12659
-
Group DETR: Fast DETR Training with Group-Wise One-to-Many Assignment 26 Jul 2022 · 2 repositories · arXiv:2207.13085
-
Multi-Attention Network for Compressed Video Referring Object Segmentation 26 Jul 2022 · 1 repository · arXiv:2207.12622
-
Remote Medication Status Prediction for Individuals with Parkinson's Disease using Time-series Data from Smartphones 26 Jul 2022 · 0 repositories · arXiv:2207.13700
-
Training Effective Neural Sentence Encoders from Automatically Mined Paraphrases 26 Jul 2022 · 1 repository · arXiv:2207.12759
-
V²L: Leveraging Vision and Vision-language Models into Large-scale Product Retrieval 26 Jul 2022 · 1 repository · arXiv:2207.12994
-
3D Siamese Transformer Network for Single Object Tracking on Point Clouds 25 Jul 2022 · 1 repository · arXiv:2207.11995
-
Behind Every Domain There is a Shift: Adapting Distortion-aware Vision Transformers for Panoramic Semantic Segmentation 25 Jul 2022 · 1 repository · arXiv:2207.11860
-
Fine-Tuning BERT for Automatic ADME Semantic Labeling in FDA Drug Labeling to Enhance Product-Specific Guidance Assessment 25 Jul 2022 · 0 repositories · arXiv:2207.12376
-
IGFormer: Interaction Graph Transformer for Skeleton-based Human Interaction Recognition 25 Jul 2022 · 0 repositories · arXiv:2207.12100
-
Is GPT-3 all you need for Visual Question Answering in Cultural Heritage? 25 Jul 2022 · 0 repositories · arXiv:2207.12101
-
Jigsaw-ViT: Learning Jigsaw Puzzles in Vision Transformer 25 Jul 2022 · 1 repository · arXiv:2207.11971
-
Reference-based Image Super-Resolution with Deformable Attention Transformer 25 Jul 2022 · 1 repository · arXiv:2207.11938
-
D3Former: Debiased Dual Distilled Transformer for Incremental Learning 25 Jul 2022 · 1 repository · arXiv:2208.00777
-
A Cognitive Study on Semantic Similarity Analysis of Large Corpora: A Transformer-based Approach 24 Jul 2022 · 0 repositories · arXiv:2207.11716
-
Affective Behaviour Analysis Using Pretrained Model with Facial Priori 24 Jul 2022 · 1 repository · arXiv:2207.11679
-
Improving Mandarin Speech Recogntion with Block-augmented Transformer 24 Jul 2022 · 2 repositories · arXiv:2207.11697
-
No More Fine-Tuning? An Experimental Evaluation of Prompt Tuning in Code Intelligence 24 Jul 2022 · 1 repository · arXiv:2207.11680Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 7 unverified (of 14 harvested samples) · 2 pointer-only (licence)
-
Online Continual Learning with Contrastive Vision Transformer 24 Jul 2022 · 0 repositories · arXiv:2207.13516
-
SAVCHOI: Detecting Suspicious Activities using Dense Video Captioning with Human Object Interactions 24 Jul 2022 · 0 repositories · arXiv:2207.11838
-
Better Reasoning Behind Classification Predictions with BERT for Fake News Detection 23 Jul 2022 · 0 repositories · arXiv:2207.11562
-
Combining Self-Training and Hybrid Architecture for Semi-supervised Abdominal Organ Segmentation 23 Jul 2022 · 2 repositories · arXiv:2207.11512
-
Comparative Validation of AI and non-AI Methods in MRI Volumetry to Diagnose Parkinsonian Syndromes 23 Jul 2022 · 0 repositories · arXiv:2207.11534
-
Context based lemmatizer for Polish language 23 Jul 2022 · 0 repositories · arXiv:2207.11565
-
High-Resolution Swin Transformer for Automatic Medical Image Segmentation 23 Jul 2022 · 1 repository · arXiv:2207.11553
-
The prediction of the quality of results in Logic Synthesis using Transformer and Graph Neural Networks 23 Jul 2022 · 0 repositories · arXiv:2207.11437
-
Applying Spatiotemporal Attention to Identify Distracted and Drowsy Driving with Vision Transformers 22 Jul 2022 · 0 repositories · arXiv:2207.12148
-
Cost Aggregation with 4D Convolutional Swin Transformer for Few-Shot Segmentation 22 Jul 2022 · 1 repository · arXiv:2207.10866Syntology official (archive's flag): 10 ran · 12 ran (of which 7 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 4 where Syntology's instrument failed) · 8 unverified (of 20 harvested samples) · 2 pointer-only (licence)
-
Hyper-Representations for Pre-Training and Transfer Learning 22 Jul 2022 · 1 repository · arXiv:2207.10951
-
Learning Generalized Non-Rigid Multimodal Biomedical Image Registration from Generic Point Set Data 22 Jul 2022 · 0 repositories · arXiv:2207.10994
-
Panoptic Scene Graph Generation 22 Jul 2022 · 1 repository · arXiv:2207.11247
-
Transformer with Implicit Edges for Particle-based Physics Simulation 22 Jul 2022 · 1 repository · arXiv:2207.10860Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Video Swin Transformers for Egocentric Video Understanding @ Ego4D Challenges 2022 22 Jul 2022 · 0 repositories · arXiv:2207.11329
-
Zero-Shot Video Captioning with Evolving Pseudo-Tokens 22 Jul 2022 · 1 repository · arXiv:2207.11100
-
Addressing Optimism Bias in Sequence Modeling for Reinforcement Learning 21 Jul 2022 · 0 repositories · arXiv:2207.10295
-
An Efficient Spatio-Temporal Pyramid Transformer for Action Detection 21 Jul 2022 · 0 repositories · arXiv:2207.10448
-
BigIssue: A Realistic Bug Localization Benchmark 21 Jul 2022 · 0 repositories · arXiv:2207.10739
-
Efficient model compression with Random Operation Access Specific Tile (ROAST) hashing 21 Jul 2022 · 1 repository · arXiv:2207.10702
-
Focused Decoding Enables 3D Anatomical Detection by Transformers 21 Jul 2022 · 1 repository · arXiv:2207.10774
-
Magic ELF: Image Deraining Meets Association Learning and Transformer 21 Jul 2022 · 1 repository · arXiv:2207.10455
-
Multi Resolution Analysis (MRA) for Approximate Self-Attention 21 Jul 2022 · 2 repositories · arXiv:2207.10284