Methods › General › Regularization › Attention Dropout › Papers, page 87
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 87 of 109: papers 8,601 to 8,700 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Efficient transfer learning for NLP with ELECTRA 6 Apr 2021 · 1 repository · arXiv:2104.02756
-
HBert + BiasCorp -- Fighting Racism on the Web 6 Apr 2021 · 0 repositories · arXiv:2104.02242
-
MuSLCAT: Multi-Scale Multi-Level Convolutional Attention Transformer for Discriminative Music Modeling on Raw Waveforms 6 Apr 2021 · 0 repositories · arXiv:2104.02309
-
COVID-19 sentiment analysis via deep learning during the rise of novel cases 5 Apr 2021 · 0 repositories · arXiv:2104.10662
-
Exploring Transformers in Emotion Recognition: a comparison of BERT, DistillBERT, RoBERTa, XLNet and ELECTRA 5 Apr 2021 · 0 repositories · arXiv:2104.02041
-
Semantic Distance: A New Metric for ASR Performance Analysis Towards Spoken Language Understanding 5 Apr 2021 · 0 repositories · arXiv:2104.02138
-
What's the best place for an AI conference, Vancouver or ______: Why completing comparative questions is difficult 5 Apr 2021 · 0 repositories · arXiv:2104.01940
-
Improving Pretrained Models for Zero-shot Multi-label Text Classification through Reinforced Label Hierarchy Reasoning 4 Apr 2021 · 1 repository · arXiv:2104.01666
-
MCL@IITK at SemEval-2021 Task 2: Multilingual and Cross-lingual Word-in-Context Disambiguation using Augmented Data, Signals, and Transformers 4 Apr 2021 · 0 repositories · arXiv:2104.01567
-
ReCAM@IITK at SemEval-2021 Task 4: BERT and ALBERT based Ensemble for Abstract Word Prediction 4 Apr 2021 · 1 repository · arXiv:2104.01563
-
Exploring the Role of BERT Token Representations to Explain Sentence Probing Results 3 Apr 2021 · 1 repository · arXiv:2104.01477
-
Unsupervised Domain Adaptation with Global and Local Graph Neural Networks in Limited Labeled Data Scenario: Application to Disaster Management 3 Apr 2021 · 0 repositories · arXiv:2104.01436
-
IITK@LCP at SemEval 2021 Task 1: Classification for Lexical Complexity Regression Task 2 Apr 2021 · 1 repository · arXiv:2104.01046
-
The Coronavirus is a Bioweapon: Analysing Coronavirus Fact-Checked Stories 2 Apr 2021 · 0 repositories · arXiv:2104.01215
-
Using GPT-2 to Create Synthetic Data to Improve the Prediction Performance of NLP Machine Learning Classification Models 2 Apr 2021 · 0 repositories · arXiv:2104.10658
-
A Dashboard for Mitigating the COVID-19 Misinfodemic 1 Apr 2021 · 0 repositories
-
Are Neural Networks Extracting Linguistic Properties or Memorizing Training Data? An Observation with a Multilingual Probe for Predicting Tense 1 Apr 2021 · 1 repository
-
BERT meets Cranfield: Uncovering the Properties of Full Ranking on Fully Labeled Data 1 Apr 2021 · 0 repositories
-
BERT Prescriptions to Avoid Unwanted Headaches: A Comparison of Transformer Architectures for Adverse Drug Event Detection 1 Apr 2021 · 1 repository
-
BERTective: Language Models and Contextual Information for Deception Detection 1 Apr 2021 · 0 repositories
-
BERxiT: Early Exiting for BERT with Better Fine-Tuning and Extension to Regression 1 Apr 2021 · 1 repository
-
Complex Question Answering on knowledge graphs using machine translation and multi-task learning 1 Apr 2021 · 0 repositories
-
Content-based Models of Quotation 1 Apr 2021 · 0 repositories
-
Detecting Scenes in Fiction: A new Segmentation Task 1 Apr 2021 · 0 repositories
-
ENPAR:Enhancing Entity and Entity Pair Representations for Joint Entity Relation Extraction 1 Apr 2021 · 1 repository
-
Evaluating language models for the retrieval and categorization of lexical collocations 1 Apr 2021 · 1 repository
-
Evaluating Neural Model Robustness for Machine Comprehension 1 Apr 2021 · 0 repositories
-
HLE-UPC at SemEval-2021 Task 5: Multi-Depth DistilBERT for Toxic Spans Detection 1 Apr 2021 · 1 repository · arXiv:2104.00639
-
How Fast can BERT Learn Simple Natural Language Inference? 1 Apr 2021 · 0 repositories
-
Keep Learning: Self-supervised Meta-learning for Learning from Inference 1 Apr 2021 · 0 repositories
-
Maximal Multiverse Learning for Promoting Cross-Task Generalization of Fine-Tuned Language Models 1 Apr 2021 · 0 repositories
-
Multilingual Entity and Relation Extraction Dataset and Model 1 Apr 2021 · 1 repository
-
Neural-Driven Search-Based Paraphrase Generation 1 Apr 2021 · 0 repositories
-
NLQuAD: A Non-Factoid Long Question Answering Data Set 1 Apr 2021 · 1 repository
-
On the (In)Effectiveness of Images for Text Classification 1 Apr 2021 · 0 repositories
-
Probing for idiomaticity in vector space models 1 Apr 2021 · 1 repository
-
Retrieval, Re-ranking and Multi-task Learning for Knowledge-Base Question Answering 1 Apr 2021 · 0 repositories
-
Russian Paraphrasers: Paraphrase with Transformers 1 Apr 2021 · 2 repositories
-
An In-depth Analysis of Passage-Level Label Transfer for Contextual Document Ranking 30 Mar 2021 · 1 repository · arXiv:2103.16669
-
Automatic Graph Partitioning for Very Large-scale Deep Learning 30 Mar 2021 · 0 repositories · arXiv:2103.16063
-
Grounding Dialogue Systems via Knowledge Graph Aware Decoding with Pre-trained Transformers 30 Mar 2021 · 1 repository · arXiv:2103.16289
-
Kaleido-BERT: Vision-Language Pre-training on Fashion Domain 30 Mar 2021 · 1 repository · arXiv:2103.16110Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Multi-Scale Vision Longformer: A New Vision Transformer for High-Resolution Image Encoding 29 Mar 2021 · 3 repositories · arXiv:2103.15358Syntology official (archive's flag): 3 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 2 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 2 pointer-only (licence)
-
Contextual Text Embeddings for Twi 29 Mar 2021 · 0 repositories · arXiv:2103.15963
-
Retraining DistilBERT for a Voice Shopping Assistant by Using Universal Dependencies 29 Mar 2021 · 0 repositories · arXiv:2103.15737
-
Whitening Sentence Representations for Better Semantics and Faster Retrieval 29 Mar 2021 · 3 repositories · arXiv:2103.15316Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples)
-
PnG BERT: Augmented BERT on Phonemes and Graphemes for Neural TTS 28 Mar 2021 · 0 repositories · arXiv:2103.15060
-
CrossViT: Cross-Attention Multi-Scale Vision Transformer for Image Classification 27 Mar 2021 · 15 repositories · arXiv:2103.14899Syntology official (archive's flag): 4 ran · 17 ran (of which 10 constructed an object rather than computing a result; 17 with no instrument failure: 0 honoured, 1 violated, 16 with no contract checked; 0 where Syntology's instrument failed) · 9 unverified (of 26 harvested samples)
-
Machine Learning Meets Natural Language Processing -- The story so far 27 Mar 2021 · 0 repositories · arXiv:2104.10213
-
Unsupervised Self-Training for Sentiment Analysis of Code-Switched Data 27 Mar 2021 · 0 repositories · arXiv:2103.14797
-
A Practical Survey on Faster and Lighter Transformers 26 Mar 2021 · 0 repositories · arXiv:2103.14636
-
BERT4SO: Neural Sentence Ordering by Fine-tuning BERT 25 Mar 2021 · 0 repositories · arXiv:2103.13584
-
Bertinho: Galician BERT Representations 25 Mar 2021 · 0 repositories · arXiv:2103.13799
-
K-XLNet: A General Method for Combining Explicit Knowledge with Language Model Pretraining 25 Mar 2021 · 0 repositories · arXiv:2104.10649
-
Predicting Directionality in Causal Relations in Text 25 Mar 2021 · 2 repositories · arXiv:2103.13606
-
Visual Grounding Strategies for Text-Only Natural Language Processing 25 Mar 2021 · 0 repositories · arXiv:2103.13942
-
Czert -- Czech BERT-like Model for Language Representation 24 Mar 2021 · 1 repository · arXiv:2103.13031
-
Low-Resource Machine Translation Training Curriculum Fit for Low-Resource Languages 24 Mar 2021 · 0 repositories · arXiv:2103.13272
-
Thinking Aloud: Dynamic Context Generation Improves Zero-Shot Reasoning Performance of GPT-2 24 Mar 2021 · 0 repositories · arXiv:2103.13033
-
Are Neural Language Models Good Plagiarists? A Benchmark for Neural Paraphrase Detection 23 Mar 2021 · 0 repositories · arXiv:2103.12450
-
Detecting Hate Speech with GPT-3 23 Mar 2021 · 2 repositories · arXiv:2103.12407
-
Repairing Pronouns in Translation with BERT-Based Post-Editing 23 Mar 2021 · 0 repositories · arXiv:2103.12838
-
The NLP Cookbook: Modern Recipes for Transformer based Deep Learning Architectures 23 Mar 2021 · 0 repositories · arXiv:2104.10640
-
TMR: Evaluating NER Recall on Tough Mentions 23 Mar 2021 · 0 repositories · arXiv:2103.12312
-
Variable Name Recovery in Decompiled Binary Code using Constrained Masked Language Modeling 23 Mar 2021 · 0 repositories · arXiv:2103.12801
-
BERT: A Review of Applications in Natural Language Processing and Understanding 22 Mar 2021 · 0 repositories · arXiv:2103.11943
-
Bridging the gap between supervised classification and unsupervised topic modelling for social-media assisted crisis management 22 Mar 2021 · 0 repositories · arXiv:2103.11835
-
PatentSBERTa: A Deep NLP based Hybrid Model for Patent Distance and Classification using Augmented SBERT 22 Mar 2021 · 2 repositories · arXiv:2103.11933
-
Identifying Machine-Paraphrased Plagiarism 22 Mar 2021 · 2 repositories · arXiv:2103.11909
-
Open Domain Question Answering over Tables via Dense Retrieval 22 Mar 2021 · 1 repository · arXiv:2103.12011
-
NameRec*: Highly Accurate and Fine-grained Person Name Recognition 21 Mar 2021 · 0 repositories · arXiv:2103.11360
-
ROSITA: Refined BERT cOmpreSsion with InTegrAted techniques 21 Mar 2021 · 1 repository · arXiv:2103.11367Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
ConViT: Improving Vision Transformers with Soft Convolutional Inductive Biases 19 Mar 2021 · 9 repositories · arXiv:2103.10697
-
Cost-effective Deployment of BERT Models in Serverless Environment 19 Mar 2021 · 0 repositories · arXiv:2103.10673
-
Let Your Heart Speak in its Mother Tongue: Multilingual Captioning of Cardiac Signals 19 Mar 2021 · 1 repository · arXiv:2103.11011
-
MuRIL: Multilingual Representations for Indian Languages 19 Mar 2021 · 1 repository · arXiv:2103.10730
-
Play the Shannon Game With Language Models: A Human-Free Approach to Summary Evaluation 19 Mar 2021 · 0 repositories · arXiv:2103.10918
-
GLM: General Language Model Pretraining with Autoregressive Blank Infilling 18 Mar 2021 · 8 repositories · arXiv:2103.10360Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Contextual Biasing of Language Models for Speech Recognition in Goal-Oriented Conversational Agents 18 Mar 2021 · 0 repositories · arXiv:2103.10325
-
GPT Understands, Too 18 Mar 2021 · 10 repositories · arXiv:2103.10385Syntology official (archive's flag): 1 ran · 3 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
Model Extraction and Adversarial Transferability, Your BERT is Vulnerable! 18 Mar 2021 · 1 repository · arXiv:2103.10013
-
Code Word Detection in Fraud Investigations using a Deep-Learning Approach 17 Mar 2021 · 0 repositories · arXiv:2103.09606
-
On the Role of Images for Analyzing Claims in Social Media 17 Mar 2021 · 1 repository · arXiv:2103.09602
-
UniParma at SemEval-2021 Task 5: Toxic Spans Detection Using CharacterBERT and Bag-of-Words Model 17 Mar 2021 · 1 repository · arXiv:2103.09645
-
KGSynNet: A Novel Entity Synonyms Discovery Framework with Knowledge Graph 16 Mar 2021 · 0 repositories · arXiv:2103.08893
-
Robustly Optimized and Distilled Training for Natural Language Understanding 16 Mar 2021 · 0 repositories · arXiv:2103.08809
-
Text Mining of Stocktwits Data for Predicting Stock Prices 13 Mar 2021 · 0 repositories · arXiv:2103.16388
-
Comparing the Performance of NLP Toolkits and Evaluation measures in Legal Tech 12 Mar 2021 · 0 repositories · arXiv:2103.11792
-
Explaining and Improving BERT Performance on Lexical Semantic Change Detection 12 Mar 2021 · 0 repositories · arXiv:2103.07259
-
Is BERT a Cross-Disciplinary Knowledge Learner? A Surprising Finding of Pre-trained Models' Transferability 12 Mar 2021 · 0 repositories · arXiv:2103.07162
-
Composite Re-Ranking for Efficient Document Search with BERT 11 Mar 2021 · 0 repositories · arXiv:2103.06499
-
Evaluation of Morphological Embeddings for the Russian Language 11 Mar 2021 · 0 repositories · arXiv:2103.06628
-
FairFil: Contrastive Neural Debiasing Method for Pretrained Text Encoders 11 Mar 2021 · 0 repositories · arXiv:2103.06413
-
Improving Bi-encoder Document Ranking Models with Two Rankers and Multi-teacher Distillation 11 Mar 2021 · 1 repository · arXiv:2103.06523
-
LightMBERT: A Simple Yet Effective Method for Multilingual BERT Distillation 11 Mar 2021 · 0 repositories · arXiv:2103.06418
-
Self-supervised Text-to-SQL Learning with Header Alignment Training 11 Mar 2021 · 0 repositories · arXiv:2103.06402
-
Towards Multi-Sense Cross-Lingual Alignment of Contextual Embeddings 11 Mar 2021 · 1 repository · arXiv:2103.06459
-
Majority Voting with Bidirectional Pre-translation For Bitext Retrieval 10 Mar 2021 · 1 repository · arXiv:2103.06369
-
CEQE: Contextualized Embeddings for Query Expansion 9 Mar 2021 · 0 repositories · arXiv:2103.05256
-
Large Pre-trained Language Models Contain Human-like Biases of What is Right and Wrong to Do 8 Mar 2021 · 1 repository · arXiv:2103.11790Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)