Methods › General › Regularization › Attention Dropout › Papers, page 99
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 99 of 109: papers 9,801 to 9,900 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Students Need More Attention: BERT-based AttentionModel for Small Data with Application to AutomaticPatient Message Triage 22 Jun 2020 · 1 repository · arXiv:2006.11991
-
MaxVA: Fast Adaptation of Step Sizes by Maximizing Observed Variance of Gradients 21 Jun 2020 · 1 repository · arXiv:2006.11918
-
Sarcasm Detection in Tweets with BERT and GloVe Embeddings 20 Jun 2020 · 0 repositories · arXiv:2006.11512
-
A Qualitative Evaluation of Language Models on Automatic Question-Answering for COVID-19 19 Jun 2020 · 1 repository · arXiv:2006.10964
-
New Vietnamese Corpus for Machine Reading Comprehension of Health News Articles 19 Jun 2020 · 0 repositories · arXiv:2006.11138
-
SqueezeBERT: What can computer vision teach NLP about efficient neural networks? 19 Jun 2020 · 6 repositories · arXiv:2006.11316Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Automatically Ranked Russian Paraphrase Corpus for Text Generation 17 Jun 2020 · 0 repositories · arXiv:2006.09719
-
Exploring the BERT Cross-Lingual Transferability: a Case Study in Reading Comprehension 17 Jun 2020 · 0 repositories
-
Tagging and parsing of multidomain collections 17 Jun 2020 · 1 repository
-
End-to-End Code Switching Language Models for Automatic Speech Recognition 16 Jun 2020 · 0 repositories · arXiv:2006.08870
-
Improving accuracy and speeding up Document Image Classification through parallel systems 16 Jun 2020 · 1 repository · arXiv:2006.09141
-
Memory-Efficient Pipeline-Parallel DNN Training 16 Jun 2020 · 1 repository · arXiv:2006.09503Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 8 unverified (of 10 harvested samples)
-
PERL: Pivot-based Domain Adaptation for Pre-trained Deep Contextualized Embedding Models 16 Jun 2020 · 1 repository · arXiv:2006.09075
-
Scalable Cross Lingual Pivots to Model Pronoun Gender for Translation 16 Jun 2020 · 0 repositories · arXiv:2006.08881
-
The SPPD System for Schema Guided Dialogue State Tracking Challenge 16 Jun 2020 · 0 repositories · arXiv:2006.09035
-
Cooking Is All About People: Comment Classification On Cookery Channels Using BERT and Classification Models (Malayalam-English Mix-Code) 15 Jun 2020 · 0 repositories · arXiv:2007.04249
-
Document Classification for COVID-19 Literature 15 Jun 2020 · 1 repository · arXiv:2006.13816
-
FinBERT: A Pretrained Language Model for Financial Communications 15 Jun 2020 · 2 repositories · arXiv:2006.08097Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
FinEst BERT and CroSloEngual BERT: less is more in multilingual models 14 Jun 2020 · 0 repositories · arXiv:2006.07890
-
Transferring Monolingual Model to Low-Resource Language: The Case of Tigrinya 13 Jun 2020 · 0 repositories · arXiv:2006.07698
-
A Monolingual Approach to Contextualized Word Embeddings for Mid-Resource Languages 11 Jun 2020 · 0 repositories · arXiv:2006.06202
-
MC-BERT: Efficient Language Pre-Training via a Meta Controller 10 Jun 2020 · 1 repository · arXiv:2006.05744Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Revisiting Few-sample BERT Fine-tuning 10 Jun 2020 · 1 repository · arXiv:2006.05987
-
Few-Shot Generative Conversational Query Rewriting 9 Jun 2020 · 1 repository · arXiv:2006.05009Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Unsupervised Paraphrase Generation using Pre-trained Language Models 9 Jun 2020 · 0 repositories · arXiv:2006.05477
-
On the Stability of Fine-tuning BERT: Misconceptions, Explanations, and Strong Baselines 8 Jun 2020 · 2 repositories · arXiv:2006.04884Syntology official (archive's flag): 8 ran · 13 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 3 honoured, 0 violated, 6 with no contract checked; 4 where Syntology's instrument failed) · 7 unverified (of 20 harvested samples) · 3 pointer-only (licence)
-
Medical Concept Normalization in User Generated Texts by Learning Target Concept Embeddings 7 Jun 2020 · 0 repositories · arXiv:2006.04014
-
Pre-training Polish Transformer-based Language Models at Scale 7 Jun 2020 · 1 repository · arXiv:2006.04229
-
Accelerating Natural Language Understanding in Task-Oriented Dialog 5 Jun 2020 · 1 repository · arXiv:2006.03701
-
DeBERTa: Decoding-enhanced BERT with Disentangled Attention 5 Jun 2020 · 14 repositories · arXiv:2006.03654Syntology official: no sample here; runs from other or unrecorded repositories · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 9 unverified (of 13 harvested samples) · 3 pointer-only (licence)
-
GMAT: Global Memory Augmentation for Transformers 5 Jun 2020 · 1 repository · arXiv:2006.03274
-
UDPipe at EvaLatin 2020: Contextualized Embeddings and Treebank Embeddings 5 Jun 2020 · 0 repositories · arXiv:2006.03687
-
The SOFC-Exp Corpus and Neural Approaches to Information Extraction in the Materials Science Domain 4 Jun 2020 · 1 repository · arXiv:2006.03039
-
Automatic Text Summarization of COVID-19 Medical Research Articles using BERT and GPT-2 3 Jun 2020 · 1 repository · arXiv:2006.01997
-
A Pairwise Probe for Understanding BERT Fine-Tuning on Machine Reading Comprehension 2 Jun 2020 · 0 repositories · arXiv:2006.01346
-
BERT Based Multilingual Machine Comprehension in English and Hindi 2 Jun 2020 · 2 repositories · arXiv:2006.01432
-
Exploring Cross-sentence Contexts for Named Entity Recognition with BERT 2 Jun 2020 · 1 repository · arXiv:2006.01563Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Position Masking for Language Models 2 Jun 2020 · 0 repositories · arXiv:2006.05676
-
Question Answering on Scholarly Knowledge Graphs 2 Jun 2020 · 0 repositories · arXiv:2006.01527
-
WikiBERT models: deep transfer learning for many languages 2 Jun 2020 · 0 repositories · arXiv:2006.01538
-
An Effective Contextual Language Modeling Framework for Speech Summarization with Augmented Features 1 Jun 2020 · 0 repositories · arXiv:2006.01189
-
BERT-based Ensembles for Modeling Disclosure and Support in Conversational Social Media Text 1 Jun 2020 · 0 repositories · arXiv:2006.01222
-
Conversational Machine Comprehension: a Literature Review 1 Jun 2020 · 0 repositories · arXiv:2006.00671
-
Emergence of Separable Manifolds in Deep Language Representations 1 Jun 2020 · 1 repository · arXiv:2006.01095
-
Étude des variations sémantiques à travers plusieurs dimensions (Studying semantic variations through several dimensions ) 1 Jun 2020 · 0 repositories
-
Introduction d'informations sémantiques dans un système de reconnaissance de la parole (Despite spectacular advances in recent years, the Automatic Speech Recognition (ASR) systems still make mistakes, especially in noisy environments) 1 Jun 2020 · 0 repositories
-
Les modèles de langue contextuels Camembert pour le français : impact de la taille et de l'hétérogénéité des données d'entrainement (C AMEM BERT Contextual Language Models for French: Impact of Training Data Size and Heterogeneity ) 1 Jun 2020 · 0 repositories
-
Qu'apporte BERT à l'analyse syntaxique en constituants discontinus ? Une suite de tests pour évaluer les prédictions de structures syntaxiques discontinues en anglais (What does BERT contribute to discontinuous constituency parsing ? A test suite to evaluate discontinuous constituency structure predictions in English) 1 Jun 2020 · 0 repositories
-
Ré-entraîner ou entraîner soi-même ? Stratégies de pré-entraînement de BERT en domaine médical (Re-train or train from scratch ? Pre-training strategies for BERT in the medical domain ) 1 Jun 2020 · 0 repositories
-
Amnesic Probing: Behavioral Explanation with Amnesic Counterfactuals 1 Jun 2020 · 0 repositories · arXiv:2006.00995
-
BPGC at SemEval-2020 Task 11: Propaganda Detection in News Articles with Multi-Granularity Knowledge Sharing and Linguistic Features based Ensemble Learning 31 May 2020 · 0 repositories · arXiv:2006.00593
-
"Judge me by my size (noun), do you?'' YodaLib: A Demographic-Aware Humor Generation Framework 31 May 2020 · 0 repositories · arXiv:2006.00578
-
LRG at SemEval-2020 Task 7: Assessing the Ability of BERT and Derivative Models to Perform Short-Edits based Humor Grading 31 May 2020 · 0 repositories · arXiv:2006.00607
-
Neural Entity Linking: A Survey of Models Based on Deep Learning 31 May 2020 · 0 repositories · arXiv:2006.00575
-
Detecting Problem Statements in Peer Assessments 30 May 2020 · 0 repositories · arXiv:2006.04532
-
A Comparative Study of Lexical Substitution Approaches based on Neural Language Models 29 May 2020 · 0 repositories · arXiv:2006.00031
-
First Neural Conjecturing Datasets and Experiments 29 May 2020 · 0 repositories · arXiv:2005.14664
-
SAFER: A Structure-free Approach for Certified Robustness to Adversarial Word Substitutions 29 May 2020 · 1 repository · arXiv:2005.14424Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Stance Prediction for Contemporary Issues: Data and Experiments 29 May 2020 · 1 repository · arXiv:2006.00052
-
Using Large Pretrained Language Models for Answering User Queries from Product Specifications 29 May 2020 · 0 repositories · arXiv:2005.14613
-
Language Models are Few-Shot Learners 28 May 2020 · 67 repositories · arXiv:2005.14165Syntology community repositories only · 45 ran (of which 0 constructed an object rather than computing a result; 40 with no instrument failure: 2 honoured, 1 violated, 37 with no contract checked; 5 where Syntology's instrument failed) · 20 unverified (of 65 harvested samples) · 7 pointer-only (licence)
-
On Incorporating Structural Information to improve Dialogue Response Generation 28 May 2020 · 1 repository · arXiv:2005.14315
-
CausaLM: Causal Model Explanation Through Counterfactual Language Models 27 May 2020 · 1 repository · arXiv:2005.13407Syntology 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
A Quantitative Survey of Communication Optimizations in Distributed Deep Learning 27 May 2020 · 1 repository · arXiv:2005.13247
-
Language Representation Models for Fine-Grained Sentiment Classification 27 May 2020 · 1 repository · arXiv:2005.13619
-
Network-to-Network Translation with Conditional Invertible Neural Networks 27 May 2020 · 1 repository · arXiv:2005.13580
-
Syntactic Structure Distillation Pretraining For Bidirectional Encoders 27 May 2020 · 0 repositories · arXiv:2005.13482
-
Transition-based Semantic Dependency Parsing with Pointer Networks 27 May 2020 · 1 repository · arXiv:2005.13344
-
A Data-driven Approach for Noise Reduction in Distantly Supervised Biomedical Relation Extraction 26 May 2020 · 1 repository · arXiv:2005.12565
-
BEEP! Korean Corpus of Online News Comments for Toxic Speech Detection 26 May 2020 · 1 repository · arXiv:2005.12503
-
BERT-XML: Large Scale Automated ICD Coding Using BERT Pretraining 26 May 2020 · 0 repositories · arXiv:2006.03685
-
Comparing BERT against traditional machine learning text classification 26 May 2020 · 0 repositories · arXiv:2005.13012
-
ParsBERT: Transformer-based Model for Persian Language Understanding 26 May 2020 · 3 repositories · arXiv:2005.12515
-
What Are People Asking About COVID-19? A Question Classification Dataset 26 May 2020 · 2 repositories · arXiv:2005.12522
-
An Audio-enriched BERT-based Framework for Spoken Multiple-choice Question Answering 25 May 2020 · 0 repositories · arXiv:2005.12142
-
Køpsala: Transition-Based Graph Parsing via Efficient Training and Effective Encoding 25 May 2020 · 1 repository · arXiv:2005.12094
-
Pointwise Paraphrase Appraisal is Potentially Problematic 25 May 2020 · 0 repositories · arXiv:2005.11996
-
Jointly Encoding Word Confusion Network and Dialogue Context with BERT for Spoken Language Understanding 24 May 2020 · 1 repository · arXiv:2005.11640
-
Comparative Study of Machine Learning Models and BERT on SQuAD 22 May 2020 · 1 repository · arXiv:2005.11313
-
L2R2: Leveraging Ranking for Abductive Reasoning 22 May 2020 · 1 repository · arXiv:2005.11223
-
Living Machines: A study of atypical animacy 22 May 2020 · 1 repository · arXiv:2005.11140
-
Med-BERT: pre-trained contextualized embeddings on large-scale structured electronic health records for disease prediction 22 May 2020 · 1 repository · arXiv:2005.12833
-
Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks 22 May 2020 · 18 repositories · arXiv:2005.11401Syntology 6 ran (of which 3 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Robust Layout-aware IE for Visually Rich Documents with Pre-trained Language Models 22 May 2020 · 0 repositories · arXiv:2005.11017
-
Text-to-Text Pre-Training for Data-to-Text Tasks 21 May 2020 · 2 repositories · arXiv:2005.10433
-
BERTweet: A pre-trained language model for English Tweets 20 May 2020 · 3 repositories · arXiv:2005.10200
-
Artificial Intelligence versus Maya Angelou: Experimental evidence that people cannot differentiate AI-generated from human-written poetry 20 May 2020 · 1 repository · arXiv:2005.09980
-
FashionBERT: Text and Image Matching with Adaptive Loss for Cross-modal Retrieval 20 May 2020 · 3 repositories · arXiv:2005.09801
-
Cross-lingual Approaches for Task-specific Dialogue Act Recognition 19 May 2020 · 0 repositories · arXiv:2005.09260
-
Sketch-BERT: Learning Sketch Bidirectional Encoder Representation from Transformers by Self-supervised Learning of Sketch Gestalt 19 May 2020 · 1 repository · arXiv:2005.09159
-
Table Search Using a Deep Contextualized Language Model 19 May 2020 · 1 repository · arXiv:2005.09207
-
Are All Languages Created Equal in Multilingual BERT? 18 May 2020 · 1 repository · arXiv:2005.09093
-
Yseop at SemEval-2020 Task 5: Cascaded BERT Language Model for Counterfactual Statement Analysis 18 May 2020 · 0 repositories · arXiv:2005.08519
-
Adversarial Training for Commonsense Inference 17 May 2020 · 1 repository · arXiv:2005.08156
-
Building a Hebrew Semantic Role Labeling Lexical Resource from Parallel Movie Subtitles 17 May 2020 · 1 repository · arXiv:2005.08206
-
Context-Based Quotation Recommendation 17 May 2020 · 0 repositories · arXiv:2005.08319
-
Cross-Lingual Low-Resource Set-to-Description Retrieval for Global E-Commerce 17 May 2020 · 1 repository · arXiv:2005.08188
-
Support-BERT: Predicting Quality of Question-Answer Pairs in MSDN using Deep Bidirectional Transformer 17 May 2020 · 0 repositories · arXiv:2005.08294
-
TaBERT: Pretraining for Joint Understanding of Textual and Tabular Data 17 May 2020 · 1 repository · arXiv:2005.08314
-
CERT: Contrastive Self-supervised Learning for Language Understanding 16 May 2020 · 0 repositories · arXiv:2005.12766