Methods › General › Regularization › Attention Dropout › Papers, page 82
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 82 of 109: papers 8,101 to 8,200 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Using Knowledge-Embedded Attention to Augment Pre-trained Language Models for Fine-Grained Emotion Recognition 31 Jul 2021 · 1 repository · arXiv:2108.00194
-
EmailSum: Abstractive Email Thread Summarization 30 Jul 2021 · 1 repository · arXiv:2107.14691
-
Perceiver IO: A General Architecture for Structured Inputs & Outputs 30 Jul 2021 · 9 repositories · arXiv:2107.14795Syntology community repositories only · 7 ran (of which 4 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples)
-
Adapting GPT, GPT-2 and BERT Language Models for Speech Recognition 29 Jul 2021 · 0 repositories · arXiv:2108.07789
-
AutoTinyBERT: Automatic Hyper-parameter Optimization for Efficient Pre-trained Language Models 29 Jul 2021 · 1 repository · arXiv:2107.13686
-
Arabic aspect sentiment polarity classification using BERT 28 Jul 2021 · 0 repositories · arXiv:2107.13290
-
gaBERT -- an Irish Language Model 27 Jul 2021 · 1 repository · arXiv:2107.12930
-
ContextNet: A Click-Through Rate Prediction Framework Using Contextual information to Refine Feature Embedding 26 Jul 2021 · 4 repositories · arXiv:2107.12025
-
Go Wider Instead of Deeper 25 Jul 2021 · 1 repository · arXiv:2107.11817
-
Context-aware Adversarial Training for Name Regularity Bias in Named Entity Recognition 24 Jul 2021 · 0 repositories · arXiv:2107.11610
-
MDQE: A More Accurate Direct Pretraining for Machine Translation Quality Estimation 24 Jul 2021 · 0 repositories · arXiv:2107.14600
-
Improving Early Sepsis Prediction with Multi Modal Learning 23 Jul 2021 · 0 repositories · arXiv:2107.11094
-
Evaluating Extractive Summarization Techniques on News Articles 22 Jul 2021 · 1 repository
-
Evaluating Extractive Summarization Techniques on News Articles 22 Jul 2021 · 1 repository
-
Evaluation of contextual embeddings on less-resourced languages 22 Jul 2021 · 0 repositories · arXiv:2107.10614
-
Semantic Text-to-Face GAN -ST^2FG 22 Jul 2021 · 0 repositories · arXiv:2107.10756
-
CausalBERT: Injecting Causal Knowledge Into Pre-trained Models with Minimal Supervision 21 Jul 2021 · 0 repositories · arXiv:2107.09852
-
DRDF: Determining the Importance of Different Multimodal Information with Dual-Router Dynamic Framework 21 Jul 2021 · 0 repositories · arXiv:2107.09909
-
Improved Text Classification via Contrastive Adversarial Training 21 Jul 2021 · 0 repositories · arXiv:2107.10137
-
Linked Data Triples Enhance Document Relevance Classification 20 Jul 2021 · 1 repository
-
Clinical Relation Extraction Using Transformer-based Models 19 Jul 2021 · 1 repository · arXiv:2107.08957
-
Aspect-based Sentiment Analysis using BERT with Disentangled Attention 18 Jul 2021 · 1 repository
-
Stock price prediction using BERT and GAN 18 Jul 2021 · 0 repositories · arXiv:2107.09055
-
A Vector-Based Approach to Few-Shot Veracity Classification for Automated Fact-Checking 17 Jul 2021 · 0 repositories
-
Neural Search: Learning Query and Product Representations in Fashion E-commerce 17 Jul 2021 · 0 repositories · arXiv:2107.08291
-
A Comparative Study of Deep Learning Classification Methods on a Small Environmental Microorganism Image Dataset (EMDS-6): from Convolutional Neural Networks to Visual Transformers 16 Jul 2021 · 0 repositories · arXiv:2107.07699
-
The Law of Large Documents: Understanding the Structure of Legal Contracts Using Visual Cues 16 Jul 2021 · 0 repositories · arXiv:2107.08128
-
AutoBERT-Zero: Evolving BERT Backbone from Scratch 15 Jul 2021 · 0 repositories · arXiv:2107.07445
-
Automatic Task Requirements Writing Evaluation via Machine Reading Comprehension 15 Jul 2021 · 1 repository · arXiv:2107.07957
-
FewCLUE: A Chinese Few-shot Learning Evaluation Benchmark 15 Jul 2021 · 1 repository · arXiv:2107.07498
-
Only Train Once: A One-Shot Neural Network Training And Pruning Framework 15 Jul 2021 · 1 repository · arXiv:2107.07467
-
Self-Supervised Contrastive Learning with Adversarial Perturbations for Defending Word Substitution-based Attacks 15 Jul 2021 · 1 repository · arXiv:2107.07610
-
Trusting RoBERTa over BERT: Insights from CheckListing the Natural Language Inference Task 15 Jul 2021 · 1 repository · arXiv:2107.07229
-
Turning Tables: Generating Examples from Semi-structured Tables for Endowing Language Models with Reasoning Skills 15 Jul 2021 · 1 repository · arXiv:2107.07261
-
BERT Fine-Tuning for Sentiment Analysis on Indonesian Mobile Apps Reviews 14 Jul 2021 · 0 repositories · arXiv:2107.06802
-
Chimera: Efficiently Training Large-Scale Neural Networks with Bidirectional Pipelines 14 Jul 2021 · 1 repository · arXiv:2107.06925
-
Indonesia's Fake News Detection using Transformer Network 14 Jul 2021 · 1 repository · arXiv:2107.06796
-
Large-Scale News Classification using BERT Language Model: Spark NLP Approach 14 Jul 2021 · 0 repositories · arXiv:2107.06785
-
Scalable Memory Protection in the PENGLAI Enclave 14 Jul 2021 · 1 repository
-
CMT: Convolutional Neural Networks Meet Vision Transformers 13 Jul 2021 · 14 repositories · arXiv:2107.06263Syntology community repositories only · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 9 harvested samples) · 3 pointer-only (licence)
-
Exploiting Network Structures to Improve Semantic Representation for the Financial Domain 13 Jul 2021 · 0 repositories · arXiv:2107.05885
-
Rating Facts under Coarse-to-fine Regimes 13 Jul 2021 · 0 repositories · arXiv:2107.06051
-
TSCAN : Dialog Structure discovery using SCAN 13 Jul 2021 · 0 repositories · arXiv:2107.06426
-
Using BERT Encoding to Tackle the Mad-lib Attack in SMS Spam Detection 13 Jul 2021 · 1 repository · arXiv:2107.06400
-
What do writing features tell us about AI papers? 13 Jul 2021 · 1 repository · arXiv:2107.06310
-
A Flexible Multi-Task Model for BERT Serving 12 Jul 2021 · 1 repository · arXiv:2107.05377
-
Asking Clarifying Questions Based on Negative Feedback in Conversational Search 12 Jul 2021 · 0 repositories · arXiv:2107.05760
-
CoBERL: Contrastive BERT for Reinforcement Learning 12 Jul 2021 · 2 repositories · arXiv:2107.05431Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
COPER: a Query-adaptable Semantics-based Search Engine for Persian COVID-19 Articles 12 Jul 2021 · 1 repository · arXiv:2107.05722
-
BERT-like Pre-training for Symbolic Piano Music Classification Tasks 12 Jul 2021 · 1 repository · arXiv:2107.05223
-
Quantifying Explainability in NLP and Analyzing Algorithms for Performance-Explainability Tradeoff 12 Jul 2021 · 1 repository · arXiv:2107.05693Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Split, embed and merge: An accurate table structure recognizer 12 Jul 2021 · 0 repositories · arXiv:2107.05214
-
Noise Stability Regularization for Improving BERT Fine-tuning 10 Jul 2021 · 0 repositories · arXiv:2107.04835
-
An Initial Investigation of Non-Native Spoken Question-Answering 9 Jul 2021 · 0 repositories · arXiv:2107.04691
-
A Review of Bangla Natural Language Processing Tasks and the Utility of Transformer Models 8 Jul 2021 · 2 repositories · arXiv:2107.03844
-
BumbleBee: A Transformer for Music 7 Jul 2021 · 0 repositories · arXiv:2107.03443
-
Can Transformer Models Measure Coherence In Text? Re-Thinking the Shuffle Test 7 Jul 2021 · 1 repository · arXiv:2107.03448
-
Efficient Transformer for Direct Speech Translation 7 Jul 2021 · 0 repositories · arXiv:2107.03069
-
Evaluating Large Language Models Trained on Code 7 Jul 2021 · 13 repositories · arXiv:2107.03374Syntology official (archive's flag): 2 ran · 26 ran (of which 0 constructed an object rather than computing a result; 24 with no instrument failure: 1 honoured, 0 violated, 23 with no contract checked; 2 where Syntology's instrument failed) · 13 unverified (of 39 harvested samples) · 4 pointer-only (licence)
-
Identifying Hijacked Reviews 7 Jul 2021 · 0 repositories · arXiv:2107.05385
-
LanguageRefer: Spatial-Language Model for 3D Visual Grounding 7 Jul 2021 · 0 repositories · arXiv:2107.03438
-
Not Quite 'Ask a Librarian': AI on the Nature, Value, and Future of LIS 7 Jul 2021 · 0 repositories · arXiv:2107.05383
-
Contradiction Detection in Persian Text 5 Jul 2021 · 0 repositories · arXiv:2107.01987
-
ERNIE 3.0: Large-scale Knowledge Enhanced Pre-training for Language Understanding and Generation 5 Jul 2021 · 2 repositories · arXiv:2107.02137
-
Experiments with adversarial attacks on text genres 5 Jul 2021 · 0 repositories · arXiv:2107.02246
-
What Helps Transformers Recognize Conversational Structure? Importance of Context, Punctuation, and Labels in Dialog Act Recognition 5 Jul 2021 · 1 repository · arXiv:2107.02294
-
KAISA: An Adaptive Second-Order Optimizer Framework for Deep Neural Networks 4 Jul 2021 · 3 repositories · arXiv:2107.01739Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
He Thinks He Knows Better than the Doctors: BERT for Event Factuality Fails on Pragmatics 2 Jul 2021 · 1 repository · arXiv:2107.00807
-
Language Identification of Hindi-English tweets using code-mixed BERT 2 Jul 2021 · 0 repositories · arXiv:2107.01202
-
Is GPT-3 Text Indistinguishable from Human Text? Scarecrow: A Framework for Scrutinizing Machine Text 2 Jul 2021 · 0 repositories · arXiv:2107.01294
-
Ultrasound Video Transformers for Cardiac Ejection Fraction Estimation 2 Jul 2021 · 1 repository · arXiv:2107.00977
-
A Primer on Pretrained Multilingual Language Models 1 Jul 2021 · 0 repositories · arXiv:2107.00676
-
AutoFormer: Searching Transformers for Visual Recognition 1 Jul 2021 · 2 repositories · arXiv:2107.00651
-
Elbert: Fast Albert with Confidence-Window Based Early Exit 1 Jul 2021 · 0 repositories · arXiv:2107.00175
-
Leveraging Domain Agnostic and Specific Knowledge for Acronym Disambiguation 1 Jul 2021 · 0 repositories · arXiv:2107.00316
-
AutoLAW: Augmented Legal Reasoning through Legal Precedent Prediction 30 Jun 2021 · 0 repositories · arXiv:2106.16034
-
Early Risk Detection of Pathological Gambling, Self-Harm and Depression Using BERT 30 Jun 2021 · 0 repositories · arXiv:2106.16175
-
Improving Factual Consistency of Abstractive Summarization on Customer Feedback 30 Jun 2021 · 0 repositories · arXiv:2106.16188
-
The MultiBERTs: BERT Reproductions for Robustness Analysis 30 Jun 2021 · 3 repositories · arXiv:2106.16163Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Hate speech detection using static BERT embeddings 29 Jun 2021 · 0 repositories · arXiv:2106.15537
-
New Arabic Medical Dataset for Diseases Classification 29 Jun 2021 · 0 repositories · arXiv:2106.15236
-
Efficient Sequence Packing without Cross-contamination: Accelerating Large Language Models without Impacting Performance 29 Jun 2021 · 1 repository · arXiv:2107.02027
-
A 3D CNN Network with BERT For Automatic COVID-19 Diagnosis From CT-Scan Images 28 Jun 2021 · 1 repository · arXiv:2106.14403
-
Current Landscape of the Russian Sentiment Corpora 28 Jun 2021 · 0 repositories · arXiv:2106.14434
-
Enhancing the Generalization for Intent Classification and Out-of-Domain Detection in SLU 28 Jun 2021 · 0 repositories · arXiv:2106.14464
-
Traditional Machine Learning and Deep Learning Models for Argumentation Mining in Russian Texts 28 Jun 2021 · 0 repositories · arXiv:2106.14438
-
What's in a Measurement? Using GPT-3 on SemEval 2021 Task 8 -- MeasEval 28 Jun 2021 · 0 repositories · arXiv:2106.14720
-
A Closer Look at How Fine-tuning Changes BERT 27 Jun 2021 · 1 repository · arXiv:2106.14282Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
AI based Presentation Creator With Customized Audio Content Delivery 27 Jun 2021 · 0 repositories · arXiv:2106.14213
-
SymbolicGPT: A Generative Transformer Model for Symbolic Regression 27 Jun 2021 · 2 repositories · arXiv:2106.14131Syntology community repositories only · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
Answering Chinese Elementary School Social Study Multiple Choice Questions 26 Jun 2021 · 0 repositories · arXiv:2107.02893
-
Benchmarking Differential Privacy and Federated Learning for BERT Models 26 Jun 2021 · 1 repository · arXiv:2106.13973
-
LNS-Madam: Low-Precision Training in Logarithmic Number System using Multiplicative Weight Update 26 Jun 2021 · 0 repositories · arXiv:2106.13914
-
SpreadsheetCoder: Formula Prediction from Semi-structured Context 26 Jun 2021 · 1 repository · arXiv:2106.15339
-
Toward Less Hidden Cost of Code Completion with Acceptance and Ranking Models 26 Jun 2021 · 0 repositories · arXiv:2106.13928
-
UMIC: An Unreferenced Metric for Image Captioning via Contrastive Learning 26 Jun 2021 · 1 repository · arXiv:2106.14019Syntology official (archive's flag): 7 ran · 7 ran (of which 4 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 14 harvested samples)
-
Adapt-and-Distill: Developing Small, Fast and Effective Pretrained Language Models for Domains 25 Jun 2021 · 0 repositories · arXiv:2106.13474
-
Learning to Sample Replacements for ELECTRA Pre-Training 25 Jun 2021 · 0 repositories · arXiv:2106.13715
-
ViTAS: Vision Transformer Architecture Search 25 Jun 2021 · 1 repository · arXiv:2106.13700
-
XL-Sum: Large-Scale Multilingual Abstractive Summarization for 44 Languages 25 Jun 2021 · 2 repositories · arXiv:2106.13822Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)