Methods › General › Regularization › Attention Dropout › Papers, page 89
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 89 of 109: papers 8,801 to 8,900 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Adversarially learning disentangled speech representations for robust multi-factor voice conversion 30 Jan 2021 · 0 repositories · arXiv:2102.00184
-
EmpathBERT: A BERT-based Framework for Demographic-aware Empathy Prediction 30 Jan 2021 · 0 repositories · arXiv:2102.00272
-
Learning From Human Correction 30 Jan 2021 · 1 repository · arXiv:2102.00225
-
ShufText: A Simple Black Box Approach to Evaluate the Fragility of Text Classification Models 30 Jan 2021 · 0 repositories · arXiv:2102.00238
-
Speech Recognition by Simply Fine-tuning BERT 30 Jan 2021 · 0 repositories · arXiv:2102.00291
-
Fine-tuning BERT-based models for Plant Health Bulletin Classification 29 Jan 2021 · 1 repository · arXiv:2102.00838
-
Synthesizing Monolingual Data for Neural Machine Translation 29 Jan 2021 · 0 repositories · arXiv:2101.12462
-
A Graph-based Relevance Matching Model for Ad-hoc Retrieval 28 Jan 2021 · 1 repository · arXiv:2101.11873
-
BERTaú: Itaú BERT for digital customer service 28 Jan 2021 · 0 repositories · arXiv:2101.12015
-
KoreALBERT: Pretraining a Lite BERT Model for Korean Language Understanding 27 Jan 2021 · 0 repositories · arXiv:2101.11363
-
On the Evolution of Syntactic Information Encoded by BERT's Contextualized Representations 27 Jan 2021 · 0 repositories · arXiv:2101.11492
-
Analyzing Zero-shot Cross-lingual Transfer in Supervised NLP Tasks 26 Jan 2021 · 0 repositories · arXiv:2101.10649
-
Attention Can Reflect Syntactic Structure (If You Let It) 26 Jan 2021 · 0 repositories · arXiv:2101.10927
-
CLiMP: A Benchmark for Chinese Language Model Evaluation 26 Jan 2021 · 0 repositories · arXiv:2101.11131
-
Deep Subjecthood: Higher-Order Grammatical Features in Multilingual BERT 26 Jan 2021 · 1 repository · arXiv:2101.11043
-
Evaluation of BERT and ALBERT Sentence Embedding Performance on Downstream NLP Tasks 26 Jan 2021 · 0 repositories · arXiv:2101.10642
-
First Align, then Predict: Understanding the Cross-Lingual Ability of Multilingual BERT 26 Jan 2021 · 1 repository · arXiv:2101.11109
-
Named Entity Recognition in the Style of Object Detection 26 Jan 2021 · 0 repositories · arXiv:2101.11122
-
Regulatory Compliance through Doc2Doc Information Retrieval: A case study in EU/UK legislation where text similarity has limitations 26 Jan 2021 · 0 repositories · arXiv:2101.10726
-
A Hybrid Approach to Measure Semantic Relatedness in Biomedical Concepts 25 Jan 2021 · 0 repositories · arXiv:2101.10196
-
Does Dialog Length matter for Next Response Selection task? An Empirical Study 24 Jan 2021 · 0 repositories · arXiv:2101.09647
-
RomeBERT: Robust Training of Multi-Exit BERT 24 Jan 2021 · 1 repository · arXiv:2101.09755
-
Stereotype and Skew: Quantifying Gender Bias in Pre-trained and Fine-tuned Language Models 24 Jan 2021 · 1 repository · arXiv:2101.09688
-
Training Multilingual Pre-trained Language Model with Byte-level Subwords 23 Jan 2021 · 1 repository · arXiv:2101.09469
-
A multi-perspective combined recall and rank framework for Chinese procedure terminology normalization 22 Jan 2021 · 0 repositories · arXiv:2101.09101
-
BERT Transformer model for Detecting Arabic GPT2 Auto-Generated Tweets 22 Jan 2021 · 0 repositories · arXiv:2101.09345
-
Drug and Disease Interpretation Learning with Biomedical Entity Representation Transformer 22 Jan 2021 · 1 repository · arXiv:2101.09311
-
Extracting Lifestyle Factors for Alzheimer's Disease from Clinical Notes Using Deep Learning with Weak Supervision 22 Jan 2021 · 0 repositories · arXiv:2101.09244
-
HASOCOne@FIRE-HASOC2020: Using BERT and Multilingual BERT models for Hate Speech Detection 22 Jan 2021 · 1 repository · arXiv:2101.09007
-
Multilingual Pre-Trained Transformers and Convolutional NN Classification Models for Technical Domain Identification 22 Jan 2021 · 0 repositories · arXiv:2101.09012
-
The heads hypothesis: A unifying statistical approach towards understanding multi-headed attention in BERT 22 Jan 2021 · 1 repository · arXiv:2101.09115
-
Evaluating Multilingual Text Encoders for Unsupervised Cross-Lingual Retrieval 21 Jan 2021 · 1 repository · arXiv:2101.08370Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Classifying Scientific Publications with BERT -- Is Self-Attention a Feature Selection Method? 20 Jan 2021 · 1 repository · arXiv:2101.08114
-
Divide and Conquer: An Ensemble Approach for Hostile Post Detection in Hindi 20 Jan 2021 · 1 repository · arXiv:2101.07973
-
Learning to Augment for Data-Scarce Domain BERT Knowledge Distillation 20 Jan 2021 · 0 repositories · arXiv:2101.08106Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 13 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
Situation and Behavior Understanding by Trope Detection on Films 19 Jan 2021 · 1 repository · arXiv:2101.07632
-
Towards Facilitating Empathic Conversations in Online Mental Health Support: A Reinforcement Learning Approach 19 Jan 2021 · 1 repository · arXiv:2101.07714
-
Automatic punctuation restoration with BERT models 18 Jan 2021 · 1 repository · arXiv:2101.07343
-
Can a Fruit Fly Learn Word Embeddings? 18 Jan 2021 · 2 repositories · arXiv:2101.06887Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Transformer-Based Models for Question Answering on COVID19 16 Jan 2021 · 0 repositories · arXiv:2101.11432
-
Grid Search Hyperparameter Benchmarking of BERT, ALBERT, and LongFormer on DuoRC 15 Jan 2021 · 0 repositories · arXiv:2101.06326
-
Hostility Detection and Covid-19 Fake News Detection in Social Media 15 Jan 2021 · 0 repositories · arXiv:2101.05953
-
KDLSQ-BERT: A Quantized Bert Combining Knowledge Distillation with Learned Step Size Quantization 15 Jan 2021 · 0 repositories · arXiv:2101.05938
-
ECOL: Early Detection of COVID Lies Using Content, Prior Knowledge and Source Information 14 Jan 2021 · 1 repository · arXiv:2101.05499
-
Persistent Anti-Muslim Bias in Large Language Models 14 Jan 2021 · 1 repository · arXiv:2101.05783
-
Transformer-based Language Model Fine-tuning Methods for COVID-19 Fake News Detection 14 Jan 2021 · 0 repositories · arXiv:2101.05509
-
WER-BERT: Automatic WER Estimation with BERT in a Balanced Ordinal Classification Paradigm 14 Jan 2021 · 0 repositories · arXiv:2101.05478
-
Experimental Evaluation of Deep Learning models for Marathi Text Classification 13 Jan 2021 · 0 repositories · arXiv:2101.04899
-
Heterogeneous Network Embedding for Deep Semantic Relevance Match in E-commerce Search 13 Jan 2021 · 0 repositories · arXiv:2101.04850
-
LaDiff ULMFiT: A Layer Differentiated training approach for ULMFiT 13 Jan 2021 · 1 repository · arXiv:2101.04965
-
Neural Contract Element Extraction Revisited: Letters from Sesame Street 12 Jan 2021 · 0 repositories · arXiv:2101.04355
-
Of Non-Linearity and Commutativity in BERT 12 Jan 2021 · 1 repository · arXiv:2101.04547
-
A More Efficient Chinese Named Entity Recognition base on BERT and Syntactic Analysis 11 Jan 2021 · 0 repositories · arXiv:2101.11423
-
AT-BERT: Adversarial Training BERT for Acronym Identification Winning Solution for SDU@AAAI-21 11 Jan 2021 · 0 repositories · arXiv:2101.03700
-
Evaluation of Deep Learning Models for Hostility Detection in Hindi Text 11 Jan 2021 · 0 repositories · arXiv:2101.04144
-
BERT & Family Eat Word Salad: Experiments with Text Understanding 10 Jan 2021 · 1 repository · arXiv:2101.03453
-
Cisco at AAAI-CAD21 shared task: Predicting Emphasis in Presentation Slides using Contextualized Embeddings 10 Jan 2021 · 1 repository · arXiv:2101.11422
-
Learning Better Sentence Representation with Syntax Information 9 Jan 2021 · 0 repositories · arXiv:2101.03343
-
Contextual Non-Local Alignment over Full-Scale Representation for Text-Based Person Search 8 Jan 2021 · 2 repositories · arXiv:2101.03036
-
Misspelling Correction with Pre-trained Contextual Language Model 8 Jan 2021 · 0 repositories · arXiv:2101.03204
-
Applying Transfer Learning for Improving Domain-Specific Search Experience Using Query to Question Similarity 7 Jan 2021 · 0 repositories · arXiv:2101.02351
-
Exploring Text-transformers in AAAI 2021 Shared Task: COVID-19 Fake News Detection in English 7 Jan 2021 · 1 repository · arXiv:2101.02359
-
Homonym Identification using BERT -- Using a Clustering Approach 7 Jan 2021 · 0 repositories · arXiv:2101.02398
-
COVID-19: Comparative Analysis of Methods for Identifying Articles Related to Therapeutics and Vaccines without Using Labeled Data 5 Jan 2021 · 0 repositories · arXiv:2101.02017
-
I-BERT: Integer-only BERT Quantization 5 Jan 2021 · 7 repositories · arXiv:2101.01321
-
Improving reference mining in patents with BERT 4 Jan 2021 · 1 repository · arXiv:2101.01039
-
A Robust and Domain-Adaptive Approach for Low-Resource Named Entity Recognition 2 Jan 2021 · 1 repository · arXiv:2101.00388
-
CDLM: Cross-Document Language Modeling 2 Jan 2021 · 2 repositories · arXiv:2101.00406
-
End-to-End Training of Neural Retrievers for Open-Domain Question Answering 2 Jan 2021 · 2 repositories · arXiv:2101.00408
-
Improving Sequence-to-Sequence Pre-training via Sequence Span Rewriting 2 Jan 2021 · 1 repository · arXiv:2101.00416
-
Lex-BERT: Enhancing BERT based NER with lexicons 2 Jan 2021 · 0 repositories · arXiv:2101.00396
-
Superbizarre Is Not Superb: Derivational Morphology Improves BERT's Interpretation of Complex Words 2 Jan 2021 · 1 repository · arXiv:2101.00403Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 2 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
What all do audio transformer models hear? Probing Acoustic Representations for Language Delivery and its Structure 2 Jan 2021 · 0 repositories · arXiv:2101.00387
-
Adding Recurrence to Pretrained Transformers 1 Jan 2021 · 0 repositories
-
BROS: A Pre-trained Language Model for Understanding Texts in Document 1 Jan 2021 · 0 repositories
-
Cluster-Former: Clustering-based Sparse Transformer for Question Answering 1 Jan 2021 · 0 repositories
-
Cluster & Tune: Enhance BERT Performance in Low Resource Text Classification 1 Jan 2021 · 0 repositories
-
Cross-Probe BERT for Efficient and Effective Cross-Modal Search 1 Jan 2021 · 0 repositories
-
DACT-BERT: Increasing the efficiency and interpretability of BERT by using adaptive computation time. 1 Jan 2021 · 0 repositories
-
Data-aware Low-Rank Compression for Large NLP Models 1 Jan 2021 · 0 repositories
-
Deep Learning Proteins using a Triplet-BERT network 1 Jan 2021 · 0 repositories
-
Domain-slot Relationship Modeling using a Pre-trained Language Encoder for Multi-Domain Dialogue State Tracking 1 Jan 2021 · 0 repositories
-
Erasure for Advancing: Dynamic Self-Supervised Learning for Commonsense Reasoning 1 Jan 2021 · 0 repositories
-
EXPLORING VULNERABILITIES OF BERT-BASED APIS 1 Jan 2021 · 0 repositories
-
How Multipurpose Are Language Models? 1 Jan 2021 · 0 repositories
-
Isotropy in the Contextual Embedding Space: Clusters and Manifolds 1 Jan 2021 · 0 repositories
-
KETG: A Knowledge Enhanced Text Generation Framework 1 Jan 2021 · 0 repositories
-
Modelling Drug-Target Binding Affinity using a BERT based Graph Neural network 1 Jan 2021 · 0 repositories
-
MULTI-SPAN QUESTION ANSWERING USING SPAN-IMAGE NETWORK 1 Jan 2021 · 0 repositories
-
On Explaining Your Explanations of BERT: An Empirical Study with Sequence Classification 1 Jan 2021 · 2 repositories · arXiv:2101.00196
-
Polyjuice: Generating Counterfactuals for Explaining, Evaluating, and Improving Models 1 Jan 2021 · 1 repository · arXiv:2101.00288
-
Post-Training Weighted Quantization of Neural Networks for Language Models 1 Jan 2021 · 0 repositories
-
Pre-training Text-to-Text Transformers to Write and Reason with Concepts 1 Jan 2021 · 0 repositories
-
Prefix-Tuning: Optimizing Continuous Prompts for Generation 1 Jan 2021 · 13 repositories · arXiv:2101.00190Syntology community repositories only · 4 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Pretrain Knowledge-Aware Language Models 1 Jan 2021 · 0 repositories
-
SkillBERT: “Skilling” the BERT to classify skills! 1 Jan 2021 · 0 repositories
-
Speeding up Deep Learning Training by Sharing Weights and Then Unsharing 1 Jan 2021 · 0 repositories
-
Syntactic Relevance XLNet Word Embedding Generation in Low-Resource Machine Translation 1 Jan 2021 · 0 repositories
-
Taking Notes on the Fly Helps Language Pre-Training 1 Jan 2021 · 0 repositories
-
Task-Agnostic and Adaptive-Size BERT Compression 1 Jan 2021 · 0 repositories