Methods › General › Attention Modules › Multi-Head Attention › Papers, page 199
Multi-Head Attention
Papers archive 2025-07-28
archive papers tagged: 24,855 · with a code link: 11,214 · where Syntology ran a sample: 3,454 (2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,454 of 24,855 tagged: 2,916 with a run with no instrument failure, 538 where every run was a failure of Syntology's instrument)
Page 199 of 249: papers 19,801 to 19,900 of 24,855, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Optimizing Latency for Online Video CaptioningUsing Audio-Visual Transformers 4 Aug 2021 · 0 repositories · arXiv:2108.02147
-
Random Offset Block Embedding Array (ROBE) for CriteoTB Benchmark MLPerf DLRM Model : 1000× Compression and 3.1× Faster Inference 4 Aug 2021 · 0 repositories · arXiv:2108.02191
-
A Dynamic Head Importance Computation Mechanism for Neural Machine Translation 3 Aug 2021 · 0 repositories · arXiv:2108.01377
-
A Study of Multilingual End-to-End Speech Recognition for Kazakh, Russian, and English 3 Aug 2021 · 1 repository · arXiv:2108.01280
-
Dynamic Feature Regularized Loss for Weakly Supervised Semantic Segmentation 3 Aug 2021 · 0 repositories · arXiv:2108.01296
-
ExBERT: An External Knowledge Enhanced BERT for Natural Language Inference 3 Aug 2021 · 0 repositories · arXiv:2108.01589
-
HTTP2vec: Embedding of HTTP Requests for Detection of Anomalous Traffic 3 Aug 2021 · 0 repositories · arXiv:2108.01763
-
Large-Scale Differentially Private BERT 3 Aug 2021 · 0 repositories · arXiv:2108.01624
-
Q-Pain: A Question Answering Dataset to Measure Social Bias in Pain Management 3 Aug 2021 · 0 repositories · arXiv:2108.01764
-
Vision Transformer with Progressive Sampling 3 Aug 2021 · 1 repository · arXiv:2108.01684Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 2 pointer-only (licence)
-
Changes in European Solidarity Before and During COVID-19: Evidence from a Large Crowd- and Expert-Annotated Twitter Dataset 2 Aug 2021 · 1 repository · arXiv:2108.01042
-
Congested Crowd Instance Localization with Dilated Convolutional Swin Transformer 2 Aug 2021 · 1 repository · arXiv:2108.00584
-
Constrained Graphic Layout Generation via Latent Optimization 2 Aug 2021 · 1 repository · arXiv:2108.00871
-
LICHEE: Improving Language Model Pre-training with Multi-grained Tokenization 2 Aug 2021 · 1 repository · arXiv:2108.00801
-
Musical Speech: A Transformer-based Composition Tool 2 Aug 2021 · 0 repositories · arXiv:2108.01043
-
Relation Aware Semi-autoregressive Semantic Parsing for NL2SQL 2 Aug 2021 · 0 repositories · arXiv:2108.00804
-
Representation learning for neural population activity with Neural Data Transformers 2 Aug 2021 · 1 repository · arXiv:2108.01210Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Self-supervised Answer Retrieval on Clinical Notes 2 Aug 2021 · 0 repositories · arXiv:2108.00775
-
Transfer Learning for Mining Feature Requests and Bug Reports from Tweets and App Store Reviews 2 Aug 2021 · 1 repository · arXiv:2108.00663
-
1213Li at SemEval-2021 Task 6: Detection of Propaganda with Multi-modal Attention and Pre-trained Models 1 Aug 2021 · 0 repositories
-
A Bidirectional Transformer Based Alignment Model for Unsupervised Word Alignment 1 Aug 2021 · 0 repositories
-
ADEPT: An Adjective-Dependent Plausibility Task 1 Aug 2021 · 0 repositories
-
Are Pretrained Convolutions Better than Pretrained Transformers? 1 Aug 2021 · 1 repository
-
AStarTwice at SemEval-2021 Task 5: Toxic Span Detection Using RoBERTa-CRF, Domain Specific Pre-Training and Self-Training 1 Aug 2021 · 0 repositories
-
AttesTable at SemEval-2021 Task 9: Extending Statement Verification with Tables for Unknown Class, and Semantic Evidence Finding 1 Aug 2021 · 1 repository
-
BennettNLP at SemEval-2021 Task 5: Toxic Spans Detection using Stacked Embedding Powered Toxic Entity Recognizer 1 Aug 2021 · 0 repositories
-
BERTAC: Enhancing Transformer-based Language Models with Adversarially Pretrained Convolutional Neural Networks 1 Aug 2021 · 1 repository
-
Best of Both Worlds: Making High Accuracy Non-incremental Transformer-based Disfluency Detection Incremental 1 Aug 2021 · 0 repositories
-
Beyond Sentence-Level End-to-End Speech Translation: Context Helps 1 Aug 2021 · 1 repository
-
Can Transformer Models Measure Coherence In Text: Re-Thinking the Shuffle Test 1 Aug 2021 · 0 repositories
-
CMTA: COVID-19 Misinformation Multilingual Analysis on Twitter 1 Aug 2021 · 0 repositories
-
Cross-lingual Evidence Improves Monolingual Fake News Detection 1 Aug 2021 · 1 repository
-
CSECU-DSG at SemEval-2021 Task 1: Fusion of Transformer Models for Lexical Complexity Prediction 1 Aug 2021 · 0 repositories
-
CSECU-DSG at SemEval-2021 Task 5: Leveraging Ensemble of Sequence Tagging Models for Toxic Spans Detection 1 Aug 2021 · 0 repositories
-
CSECU-DSG at SemEval-2021 Task 6: Orchestrating Multimodal Neural Architectures for Identifying Persuasion Techniques in Texts and Images 1 Aug 2021 · 0 repositories
-
CSECU-DSG at SemEval-2021 Task 7: Detecting and Rating Humor and Offense Employing Transformers 1 Aug 2021 · 0 repositories
-
CTFN: Hierarchical Learning for Multimodal Sentiment Analysis Using Coupled-Translation Fusion Network 1 Aug 2021 · 1 repository
-
Deep Differential Amplifier for Extractive Summarization 1 Aug 2021 · 0 repositories
-
DeepBlueAI at SemEval-2021 Task 7: Detecting and Rating Humor and Offense with Stacking Diverse Language Model-Based Methods 1 Aug 2021 · 0 repositories
-
DLJUST at SemEval-2021 Task 7: Hahackathon: Linking Humor and Offense 1 Aug 2021 · 0 repositories
-
Early Detection of Sexual Predators in Chats 1 Aug 2021 · 1 repository
-
Ecco: An Open Source Library for the Explainability of Transformer Language Models 1 Aug 2021 · 1 repository
-
eMLM: A New Pre-training Objective for Emotion Related Tasks 1 Aug 2021 · 1 repository
-
Employing Argumentation Knowledge Graphs for Neural Argument Generation 1 Aug 2021 · 1 repository
-
EndTimes at SemEval-2021 Task 7: Detecting and Rating Humor and Offense with BERT and Ensembles 1 Aug 2021 · 0 repositories
-
ES-JUST at SemEval-2021 Task 7: Detecting and Rating Humor and Offensive Text Using Deep Learning 1 Aug 2021 · 0 repositories
-
Explaining Contextualization in Language Models using Visual Analytics 1 Aug 2021 · 0 repositories
-
Explanations for CommonsenseQA: New Dataset and Models 1 Aug 2021 · 0 repositories
-
Exploring Listwise Evidence Reasoning with T5 for Fact Verification 1 Aug 2021 · 0 repositories
-
Fast and Accurate Neural Machine Translation with Translation Memory 1 Aug 2021 · 0 repositories
-
GHOST at SemEval-2021 Task 5: Is explanation all you need? 1 Aug 2021 · 0 repositories
-
GhostBERT: Generate More Features with Cheap Operations for BERT 1 Aug 2021 · 0 repositories
-
Grenzlinie at SemEval-2021 Task 7: Detecting and Rating Humor and Offense 1 Aug 2021 · 0 repositories
-
Gulu at SemEval-2021 Task 7: Detecting and Rating Humor and Offense 1 Aug 2021 · 0 repositories
-
GX at SemEval-2021 Task 2: BERT with Lemma Information for MCL-WiC Task 1 Aug 2021 · 1 repository
-
HamiltonDinggg at SemEval-2021 Task 5: Investigating Toxic Span Detection using RoBERTa Pre-training 1 Aug 2021 · 0 repositories
-
How effective is BERT without word ordering? Implications for language understanding and data privacy 1 Aug 2021 · 0 repositories
-
How Many Layers and Why? An Analysis of the Model Depth in Transformers 1 Aug 2021 · 0 repositories
-
hub at SemEval-2021 Task 1: Fusion of Sentence and Word Frequency to Predict Lexical Complexity 1 Aug 2021 · 0 repositories
-
hub at SemEval-2021 Task 2: Word Meaning Similarity Prediction Model Based on RoBERTa and Word Frequency 1 Aug 2021 · 0 repositories
-
hub at SemEval-2021 Task 7: Fusion of ALBERT and Word Frequency Information Detecting and Rating Humor and Offense 1 Aug 2021 · 0 repositories
-
IITK@LCP at SemEval-2021 Task 1: Classification for Lexical Complexity Regression Task 1 Aug 2021 · 0 repositories
-
Issues with Entailment-based Zero-shot Text Classification 1 Aug 2021 · 1 repository
-
JCT at SemEval-2021 Task 1: Context-aware Representation for Lexical Complexity Prediction 1 Aug 2021 · 0 repositories
-
JUST-BLUE at SemEval-2021 Task 1: Predicting Lexical Complexity using BERT and RoBERTa Pre-trained Language Models 1 Aug 2021 · 0 repositories
-
KuiLeiXi: a Chinese Open-Ended Text Adventure Game 1 Aug 2021 · 0 repositories
-
LASOR: Learning Accurate 3D Human Pose and Shape Via Synthetic Occlusion-Aware Data and Neural Mesh Rendering 1 Aug 2021 · 1 repository · arXiv:2108.00351Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
LeCun at SemEval-2021 Task 6: Detecting Persuasion Techniques in Text Using Ensembled Pretrained Transformers and Data Augmentation 1 Aug 2021 · 0 repositories
-
LeeBERT: Learned Early Exit for BERT with cross-level optimization 1 Aug 2021 · 0 repositories
-
LIORI at SemEval-2021 Task 8: Ask Transformer for measurements 1 Aug 2021 · 0 repositories
-
Lotus at SemEval-2021 Task 2: Combination of BERT and Paraphrasing for English Word Sense Disambiguation 1 Aug 2021 · 0 repositories
-
Measure and Evaluation of Semantic Divergence across Two Languages 1 Aug 2021 · 0 repositories
-
Measuring and Improving BERT's Mathematical Abilities by Predicting the Order of Reasoning. 1 Aug 2021 · 0 repositories
-
MedAI at SemEval-2021 Task 5: Start-to-end Tagging Framework for Toxic Spans Detection 1 Aug 2021 · 0 repositories
-
MinD at SemEval-2021 Task 6: Propaganda Detection using Transfer Learning and Multimodal Fusion 1 Aug 2021 · 0 repositories
-
Modeling Task-Aware MIMO Cardinality for Efficient Multilingual Neural Machine Translation 1 Aug 2021 · 0 repositories
-
More than Text: Multi-modal Chinese Word Segmentation 1 Aug 2021 · 1 repository
-
Multi-Head Highly Parallelized LSTM Decoder for Neural Machine Translation 1 Aug 2021 · 0 repositories
-
MVP-BERT: Multi-Vocab Pre-training for Chinese BERT 1 Aug 2021 · 0 repositories
-
Neural Metaphor Detection with Visibility Embeddings 1 Aug 2021 · 0 repositories
-
NLPIITR at SemEval-2021 Task 6: RoBERTa Model with Data Augmentation for Persuasion Techniques Detection 1 Aug 2021 · 0 repositories
-
NLyticsFKIE at SemEval-2021 Task 6: Detection of Persuasion Techniques In Texts And Images 1 Aug 2021 · 0 repositories
-
nmT5 - Is parallel data still relevant for pre-training massively multilingual language models? 1 Aug 2021 · 0 repositories
-
On the differences between BERT and MT encoder spaces and how to address them in translation tasks 1 Aug 2021 · 0 repositories
-
PAW at SemEval-2021 Task 2: Multilingual and Cross-lingual Word-in-Context Disambiguation : Exploring Cross Lingual Transfer, Augmentations and Adversarial Training 1 Aug 2021 · 0 repositories
-
PINGAN Omini-Sinitic at SemEval-2021 Task 4:Reading Comprehension of Abstract Meaning 1 Aug 2021 · 0 repositories
-
PLOME: Pre-training with Misspelled Knowledge for Chinese Spelling Correction 1 Aug 2021 · 1 repository
-
Point, Disambiguate and Copy: Incorporating Bilingual Dictionaries for Neural Machine Translation 1 Aug 2021 · 0 repositories
-
PRAL: A Tailored Pre-Training Model for Task-Oriented Dialog Generation 1 Aug 2021 · 0 repositories
-
Predicting pragmatic discourse features in the language of adults with autism spectrum disorder 1 Aug 2021 · 0 repositories
-
RAW-C: Relatedness of Ambiguous Words in Context (A New Lexical Resource for English) 1 Aug 2021 · 0 repositories
-
RG PA at SemEval-2021 Task 1: A Contextual Attention-based Model with RoBERTa for Lexical Complexity Prediction 1 Aug 2021 · 0 repositories
-
SarcasmDet at SemEval-2021 Task 7: Detect Humor and Offensive based on Demographic Factors using RoBERTa Pre-trained Model 1 Aug 2021 · 0 repositories
-
Sefamerve ARGE at SemEval-2021 Task 5: Toxic Spans Detection Using Segmentation Based 1-D Convolutional Neural Network Model 1 Aug 2021 · 1 repository
-
SkoltechNLP at SemEval-2021 Task 5: Leveraging Sentence-level Pre-training for Toxic Span Detection 1 Aug 2021 · 0 repositories
-
Stanford MLab at SemEval-2021 Task 8: 48 Hours Is All You Need 1 Aug 2021 · 0 repositories
-
Stretch-VST: Getting Flexible With Visual Stories 1 Aug 2021 · 0 repositories
-
Synchronous Syntactic Attention for Transformer Neural Machine Translation 1 Aug 2021 · 0 repositories
-
Taming Pre-trained Language Models with N-gram Representations for Low-Resource Domain Adaptation 1 Aug 2021 · 1 repository
-
Team_KGP at SemEval-2021 Task 7: A Deep Neural System to Detect Humor and Offense with Their Ratings in the Text Data 1 Aug 2021 · 0 repositories