Methods › General › Attention Modules › Multi-Head Attention › Papers, page 184
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 184 of 249: papers 18,301 to 18,400 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Multi-Modal Dynamic Graph Transformer for Visual Grounding 1 Jan 2022 · 1 repository
-
Neural Window Fully-Connected CRFs for Monocular Depth Estimation 1 Jan 2022 · 0 repositories
-
PatchTrack: Multiple Object Tracking Using Frame Patches 1 Jan 2022 · 0 repositories · arXiv:2201.00080
-
Recurring the Transformer for Video Action Recognition 1 Jan 2022 · 0 repositories
-
SpaceEdit: Learning a Unified Editing Space for Open-Domain Image Color Editing 1 Jan 2022 · 0 repositories
-
Tencent-MVSE: A Large-Scale Benchmark Dataset for Multi-Modal Video Similarity Evaluation 1 Jan 2022 · 0 repositories
-
The GatedTabTransformer. An enhanced deep learning architecture for tabular modeling 1 Jan 2022 · 2 repositories · arXiv:2201.00199
-
Training Object Detectors From Scratch: An Empirical Study in the Era of Vision Transformer 1 Jan 2022 · 0 repositories
-
Transformer Based Line Segment Classifier With Image Context for Real-Time Vanishing Point Detection in Manhattan World 1 Jan 2022 · 0 repositories
-
Uncertainty-Guided Probabilistic Transformer for Complex Action Recognition 1 Jan 2022 · 0 repositories
-
A Neural Network Solves, Explains, and Generates University Math Problems by Program Synthesis and Few-Shot Learning at Human Level 31 Dec 2021 · 1 repository · arXiv:2112.15594Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Clustering Vietnamese Conversations From Facebook Page To Build Training Dataset For Chatbot 31 Dec 2021 · 1 repository · arXiv:2112.15338
-
CSformer: Bridging Convolution and Transformer for Compressive Sensing 31 Dec 2021 · 1 repository · arXiv:2112.15299
-
Multi-Dimensional Model Compression of Vision Transformer 31 Dec 2021 · 1 repository · arXiv:2201.00043
-
OpenQA: Hybrid QA System Relying on Structured Knowledge Base as well as Non-structured Data 31 Dec 2021 · 0 repositories · arXiv:2112.15356
-
Scene-Adaptive Attention Network for Crowd Counting 31 Dec 2021 · 0 repositories · arXiv:2112.15509
-
Transformer Embeddings of Irregularly Spaced Events and Their Participants 31 Dec 2021 · 2 repositories · arXiv:2201.00044
-
ViNMT: Neural Machine Translation Toolkit 31 Dec 2021 · 1 repository · arXiv:2112.15272
-
A Lightweight and Accurate Spatial-Temporal Transformer for Traffic Forecasting 30 Dec 2021 · 1 repository · arXiv:2201.00008
-
Automatic Mixed-Precision Quantization Search of BERT 30 Dec 2021 · 0 repositories · arXiv:2112.14938
-
ChunkFormer: Learning Long Time Series with Multi-stage Chunked Transformer 30 Dec 2021 · 0 repositories · arXiv:2112.15087
-
Persformer: A Transformer Architecture for Topological Machine Learning 30 Dec 2021 · 1 repository · arXiv:2112.15210
-
THE Benchmark: Transferable Representation Learning for Monocular Height Estimation 30 Dec 2021 · 0 repositories · arXiv:2112.14985
-
EvoMoE: An Evolutional Mixture-of-Experts Training Framework via Dense-To-Sparse Gate 29 Dec 2021 · 2 repositories · arXiv:2112.14397Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Learning Spatially-Adaptive Squeeze-Excitation Networks for Image Synthesis and Image Recognition 29 Dec 2021 · 1 repository · arXiv:2112.14804
-
Temporal Attention Augmented Transformer Hawkes Process 29 Dec 2021 · 0 repositories · arXiv:2112.14472
-
Universal Transformer Hawkes Process with Adaptive Recursive Iteration 29 Dec 2021 · 0 repositories · arXiv:2112.14479
-
APRIL: Finding the Achilles' Heel on Privacy for Vision Transformers 28 Dec 2021 · 1 repository · arXiv:2112.14087
-
Extended Self-Critical Pipeline for Transforming Videos to Text (TRECVID-VTT Task 2021) -- Team: MMCUniAugsburg 28 Dec 2021 · 0 repositories · arXiv:2112.14100
-
Pale Transformer: A General Vision Transformer Backbone with Pale-Shaped Attention 28 Dec 2021 · 2 repositories · arXiv:2112.14000Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Synchronized Audio-Visual Frames with Fractional Positional Encoding for Transformers in Video-to-Text Translation 28 Dec 2021 · 0 repositories · arXiv:2112.14088
-
The University of Texas at Dallas HLTRI's Participation in EPIC-QA: Searching for Entailed Questions Revealing Novel Answer Nuggets 28 Dec 2021 · 0 repositories · arXiv:2112.13946
-
"A Passage to India": Pre-trained Word Embeddings for Indian Languages 27 Dec 2021 · 0 repositories · arXiv:2112.13800
-
Contextual Sentence Analysis for the Sentiment Prediction on Financial Data 27 Dec 2021 · 0 repositories · arXiv:2112.13790
-
Event-based clinical findings extraction from radiology reports with pre-trained language model 27 Dec 2021 · 1 repository · arXiv:2112.13512
-
HeteroQA: Learning towards Question-and-Answering through Multiple Information Sources via Heterogeneous Graph Modeling 27 Dec 2021 · 1 repository · arXiv:2112.13597
-
Learning Generative Vision Transformer with Energy-Based Latent Space for Saliency Prediction 27 Dec 2021 · 0 repositories · arXiv:2112.13528
-
Learning Robust and Lightweight Model through Separable Structured Transformations 27 Dec 2021 · 0 repositories · arXiv:2112.13551
-
Mind the Gap: Cross-Lingual Information Retrieval with Hierarchical Knowledge Enhancement 27 Dec 2021 · 0 repositories · arXiv:2112.13510
-
MSHT: Multi-stage Hybrid Transformer for the ROSE Image Analysis of Pancreatic Cancer 27 Dec 2021 · 1 repository · arXiv:2112.13513
-
Multi-Image Visual Question Answering 27 Dec 2021 · 1 repository · arXiv:2112.13706
-
Secondary Use of Clinical Problem List Entries for Neural Network-Based Disease Code Assignment 27 Dec 2021 · 0 repositories · arXiv:2112.13756
-
SPViT: Enabling Faster Vision Transformers via Soft Token Pruning 27 Dec 2021 · 1 repository · arXiv:2112.13890Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Evaluating Contextual Embeddings and their Extraction Layers for Depression Assessment 27 Dec 2021 · 0 repositories · arXiv:2112.13795
-
Video Joint Modelling Based on Hierarchical Transformer for Co-summarization 27 Dec 2021 · 2 repositories · arXiv:2112.13478
-
ViR:the Vision Reservoir 27 Dec 2021 · 0 repositories · arXiv:2112.13545
-
Vision Transformer for Small-Size Datasets 27 Dec 2021 · 5 repositories · arXiv:2112.13492Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
An Ensemble of Pre-trained Transformer Models For Imbalanced Multiclass Malware Classification 25 Dec 2021 · 1 repository · arXiv:2112.13236
-
CABACE: Injecting Character Sequence Information and Domain Knowledge for Enhanced Acronym and Long-Form Extraction 25 Dec 2021 · 1 repository · arXiv:2112.13237
-
Combining Improvements for Exploiting Dependency Trees in Neural Semantic Parsing 25 Dec 2021 · 0 repositories · arXiv:2112.13179
-
Deeper Clinical Document Understanding Using Relation Extraction 25 Dec 2021 · 1 repository · arXiv:2112.13259
-
Raw Produce Quality Detection with Shifted Window Self-Attention 24 Dec 2021 · 0 repositories · arXiv:2112.13845
-
SimViT: Exploring a Simple Vision Transformer with sliding windows 24 Dec 2021 · 2 repositories · arXiv:2112.13085
-
Distilling the Knowledge of Romanian BERTs Using Multiple Teachers 23 Dec 2021 · 1 repository · arXiv:2112.12650
-
ELSA: Enhanced Local Self-Attention for Vision Transformer 23 Dec 2021 · 1 repository · arXiv:2112.12786Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
ERNIE 3.0 Titan: Exploring Larger-scale Knowledge Enhanced Pre-training for Language Understanding and Generation 23 Dec 2021 · 3 repositories · arXiv:2112.12731
-
LaTr: Layout-Aware Transformer for Scene-Text VQA 23 Dec 2021 · 1 repository · arXiv:2112.12494Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples)
-
S+PAGE: A Speaker and Position-Aware Graph Neural Network Model for Emotion Recognition in Conversation 23 Dec 2021 · 0 repositories · arXiv:2112.12389
-
SeMask: Semantically Masked Transformers for Semantic Segmentation 23 Dec 2021 · 1 repository · arXiv:2112.12782
-
Adaptive Beam Search to Enhance On-device Abstractive Summarization 22 Dec 2021 · 0 repositories · arXiv:2201.02739
-
Comprehensive Visual Question Answering on Point Clouds through Compositional Scene Manipulation 22 Dec 2021 · 1 repository · arXiv:2112.11691Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Consistency and Coherence from Points of Contextual Similarity 22 Dec 2021 · 0 repositories · arXiv:2112.11638
-
DA-FDFtNet: Dual Attention Fake Detection Fine-tuning Network to Detect Various AI-Generated Fake Images 22 Dec 2021 · 0 repositories · arXiv:2112.12001
-
Diformer: Directional Transformer for Neural Machine Translation 22 Dec 2021 · 0 repositories · arXiv:2112.11632
-
Contrast and Generation Make BART a Good Dialogue Emotion Recognizer 21 Dec 2021 · 1 repository · arXiv:2112.11202
-
DB-BERT: a Database Tuning Tool that "Reads the Manual" 21 Dec 2021 · 0 repositories · arXiv:2112.10925
-
iSegFormer: Interactive Segmentation via Transformers with Application to 3D Knee MR Images 21 Dec 2021 · 1 repository · arXiv:2112.11325
-
Learned Queries for Efficient Local Attention 21 Dec 2021 · 1 repository · arXiv:2112.11435Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
MIA-Former: Efficient and Robust Vision Transformers via Multi-grained Input-Adaptation 21 Dec 2021 · 0 repositories · arXiv:2112.11542
-
MPViT: Multi-Path Vision Transformer for Dense Prediction 21 Dec 2021 · 3 repositories · arXiv:2112.11010
-
Predicting Job Titles from Job Descriptions with Multi-label Text Classification 21 Dec 2021 · 1 repository · arXiv:2112.11052
-
SOIT: Segmenting Objects with Instance-Aware Transformers 21 Dec 2021 · 1 repository · arXiv:2112.11037
-
Article Reranking by Memory-Enhanced Key Sentence Matching for Detecting Previously Fact-Checked Claims 20 Dec 2021 · 1 repository · arXiv:2112.10322
-
Diaformer: Automatic Diagnosis via Symptoms Sequence Generation 20 Dec 2021 · 1 repository · arXiv:2112.10433
-
Few-shot Learning with Multilingual Language Models 20 Dec 2021 · 2 repositories · arXiv:2112.10668
-
Lite Vision Transformer with Enhanced Self-Attention 20 Dec 2021 · 1 repository · arXiv:2112.10809
-
Multi-Singer: Fast Multi-Singer Singing Voice Vocoder With A Large-Scale Corpus 20 Dec 2021 · 2 repositories · arXiv:2112.10358Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Self-attention Presents Low-dimensional Knowledge Graph Embeddings for Link Prediction 20 Dec 2021 · 1 repository · arXiv:2112.10644
-
StyleSwin: Transformer-based GAN for High-resolution Image Generation 20 Dec 2021 · 1 repository · arXiv:2112.10762Syntology official (archive's flag): 2 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
Training dataset and dictionary sizes matter in BERT models: the case of Baltic languages 20 Dec 2021 · 0 repositories · arXiv:2112.10553
-
Analysis and Mitigation of Dataset Artifacts in OpenAI GPT-3 19 Dec 2021 · 0 repositories
-
Data Augmentation for Mental Health Classification on Social Media 19 Dec 2021 · 0 repositories · arXiv:2112.10064
-
SSDNet: State Space Decomposition Neural Network for Time Series Forecasting 19 Dec 2021 · 1 repository · arXiv:2112.10251
-
Task-Oriented Multi-User Semantic Communications 19 Dec 2021 · 0 repositories · arXiv:2112.10255
-
Exploiting Long-Term Dependencies for Generating Dynamic Scene Graphs 18 Dec 2021 · 1 repository · arXiv:2112.09828
-
Leveraging Transformers for Hate Speech Detection in Conversational Code-Mixed Tweets 18 Dec 2021 · 0 repositories · arXiv:2112.09986
-
Zero-shot and Few-shot Learning with Knowledge Graphs: A Comprehensive Survey 18 Dec 2021 · 0 repositories · arXiv:2112.10006
-
Syntactic-GCN Bert based Chinese Event Extraction 18 Dec 2021 · 0 repositories · arXiv:2112.09939
-
A High-Precision Health-relatedness Score for Phrases to Mine Cause-Effect Statements from the Web 17 Dec 2021 · 0 repositories
-
A Simple Single-Scale Vision Transformer for Object Localization and Instance Segmentation 17 Dec 2021 · 3 repositories · arXiv:2112.09747
-
Can Machine Learning Tools Support the Identification of Sustainable Design Leads From Product Reviews? Opportunities and Challenges 17 Dec 2021 · 0 repositories · arXiv:2112.09391
-
Challenging America: Modeling language in longer time scales 17 Dec 2021 · 0 repositories
-
Efficient Visual Tracking with Exemplar Transformers 17 Dec 2021 · 2 repositories · arXiv:2112.09686
-
Explain, Edit, and Understand: Rethinking User Study Design for Evaluating Model Explanations 17 Dec 2021 · 1 repository · arXiv:2112.09669
-
Full Transformer Framework for Robust Point Cloud Registration with Deep Information Interaction 17 Dec 2021 · 1 repository · arXiv:2112.09385
-
Incorporate Dependency Relation Knowledge into Transformer Block for Multi-turn Dialogue Generation 17 Dec 2021 · 0 repositories
-
Joint Chinese Word Segmentation and Part-of-speech Tagging via Two-stage Span Labeling 17 Dec 2021 · 0 repositories · arXiv:2112.09488
-
Learning to Win Lottery Tickets in BERT Transfer via Task-agnostic Mask Training 17 Dec 2021 · 0 repositories
-
MOVER: Mask, Over-generate and Rank for Hyperbole Generation 17 Dec 2021 · 0 repositories
-
Rank4Class: A Ranking Formulation for Multiclass Classification 17 Dec 2021 · 0 repositories · arXiv:2112.09727