Methods › General › Regularization › Attention Dropout › Papers, page 65
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 65 of 109: papers 6,401 to 6,500 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Supervised Contrastive Learning as Multi-Objective Optimization for Fine-Tuning Large Pre-trained Language Models 28 Sep 2022 · 0 repositories · arXiv:2209.14161
-
Who is GPT-3? An Exploration of Personality, Values and Demographics 28 Sep 2022 · 1 repository · arXiv:2209.14338
-
YATO: Yet Another deep learning based Text analysis Open toolkit 28 Sep 2022 · 1 repository · arXiv:2209.13877
-
How GPT-3 responds to different publics on climate change and Black Lives Matter: A critical appraisal of equity in conversational AI 27 Sep 2022 · 0 repositories · arXiv:2209.13627
-
Extractive Question Answering on Queries in Hindi and Tamil 27 Sep 2022 · 0 repositories · arXiv:2210.06356
-
Outlier Suppression: Pushing the Limit of Low-bit Transformer Language Models 27 Sep 2022 · 1 repository · arXiv:2209.13325Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
WikiDes: A Wikipedia-Based Dataset for Generating Short Descriptions from Paragraphs 27 Sep 2022 · 1 repository · arXiv:2209.13101
-
Do ever larger octopi still amplify reporting biases? Evidence from judgments of typical colour 26 Sep 2022 · 0 repositories · arXiv:2209.12786
-
News Summarization and Evaluation in the Era of GPT-3 26 Sep 2022 · 1 repository · arXiv:2209.12356
-
Towards Simple and Efficient Task-Adaptive Pre-training for Text Classification 26 Sep 2022 · 0 repositories · arXiv:2209.12943
-
Application of Deep Learning in Generating Structured Radiology Reports: A Transformer-Based Technique 25 Sep 2022 · 1 repository · arXiv:2209.12177
-
SpeedLimit: Neural Architecture Search for Quantized Transformer Models 25 Sep 2022 · 0 repositories · arXiv:2209.12127
-
Sentiment Analysis on Inflation after Covid-19 25 Sep 2022 · 0 repositories · arXiv:2209.14737
-
Can Transformer Models Effectively Detect Software Aspects in StackOverflow Discussion? 24 Sep 2022 · 0 repositories · arXiv:2209.12065
-
Learning Chess With Language Models and Transformers 24 Sep 2022 · 0 repositories · arXiv:2209.11902
-
Moral Mimicry: Large Language Models Produce Moral Rationalizations Tailored to Political Identity 24 Sep 2022 · 0 repositories · arXiv:2209.12106
-
ET5: A Novel End-to-end Framework for Conversational Machine Reading Comprehension 23 Sep 2022 · 1 repository · arXiv:2209.11484
-
IDEA: Interactive DoublE Attentions from Label Embedding for Text Classification 23 Sep 2022 · 0 repositories · arXiv:2209.11407
-
A Case Report On The "A.I. Locked-In Problem": social concerns with modern NLP 22 Sep 2022 · 0 repositories · arXiv:2209.12687
-
Adaptation of domain-specific transformer models with text oversampling for sentiment analysis of social media posts on Covid-19 vaccines 22 Sep 2022 · 1 repository · arXiv:2209.10966
-
AIR-JPMC@SMM4H'22: Classifying Self-Reported Intimate Partner Violence in Tweets with Multiple BERT-based Models 22 Sep 2022 · 0 repositories · arXiv:2209.10763
-
DFX: A Low-latency Multi-FPGA Appliance for Accelerating Transformer-based Text Generation 22 Sep 2022 · 0 repositories · arXiv:2209.10797
-
XF2T: Cross-lingual Fact-to-Text Generation for Low-Resource Languages 22 Sep 2022 · 0 repositories · arXiv:2209.11252
-
Bias at a Second Glance: A Deep Dive into Bias for German Educational Peer-Review Data Modeling 21 Sep 2022 · 2 repositories · arXiv:2209.10335Syntology official: harvested, nothing ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
CAE: Mechanism to Diminish the Class Imbalanced in SLU Slot Filling Task 21 Sep 2022 · 1 repository
-
Representing Affect Information in Word Embeddings 21 Sep 2022 · 0 repositories · arXiv:2209.10583
-
Subject Verb Agreement Error Patterns in Meaningless Sentences: Humans vs. BERT 21 Sep 2022 · 0 repositories · arXiv:2209.10538
-
T5QL: Taming language models for SQL generation 21 Sep 2022 · 0 repositories · arXiv:2209.10254
-
Text Revealer: Private Text Reconstruction via Model Inversion Attacks against Transformers 21 Sep 2022 · 0 repositories · arXiv:2209.10505
-
Towards Fine-tuning Pre-trained Language Models with Integer Forward and Backward Propagation 20 Sep 2022 · 0 repositories · arXiv:2209.09815
-
One-to-Many Semantic Communication Systems: Design, Implementation, Performance Evaluation 20 Sep 2022 · 0 repositories · arXiv:2209.09425
-
Unsupervised Early Exit in DNNs with Multiple Exits 20 Sep 2022 · 1 repository · arXiv:2209.09480
-
Meta-Adapters: Parameter Efficient Few-shot Fine-tuning through Meta-Learning 19 Sep 2022 · 1 repository
-
Will It Blend? Mixing Training Paradigms & Prompting for Argument Quality Prediction 19 Sep 2022 · 0 repositories · arXiv:2209.08966
-
Detecting Generated Scientific Papers using an Ensemble of Transformer Models 17 Sep 2022 · 1 repository · arXiv:2209.08283
-
CodeQueries: A Dataset of Semantic Queries over Code 17 Sep 2022 · 1 repository · arXiv:2209.08372
-
Changing the Representation: Examining Language Representation for Neural Sign Language Production 16 Sep 2022 · 0 repositories · arXiv:2210.06312
-
Psychologically-informed chain-of-thought prompts for metaphor understanding in large language models 16 Sep 2022 · 1 repository · arXiv:2209.08141Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Text and Patterns: For Effective Chain of Thought, It Takes Two to Tango 16 Sep 2022 · 0 repositories · arXiv:2209.07686
-
Machine Reading, Fast and Slow: When Do Models "Understand" Language? 15 Sep 2022 · 0 repositories · arXiv:2209.07430
-
uChecker: Masked Pretrained Language Models as Unsupervised Chinese Spelling Checkers 15 Sep 2022 · 0 repositories · arXiv:2209.07068
-
Automated Fidelity Assessment for Strategy Training in Inpatient Rehabilitation using Natural Language Processing 14 Sep 2022 · 0 repositories · arXiv:2209.06727
-
BERT-based Ensemble Approaches for Hate Speech Detection 14 Sep 2022 · 0 repositories · arXiv:2209.06505
-
Efficient Quantized Sparse Matrix Operations on Tensor Cores 14 Sep 2022 · 1 repository · arXiv:2209.06979
-
Out of One, Many: Using Language Models to Simulate Human Samples 14 Sep 2022 · 0 repositories · arXiv:2209.06899
-
Pre-training for Information Retrieval: Are Hyperlinks Fully Explored? 14 Sep 2022 · 0 repositories · arXiv:2209.06583
-
CNN-Trans-Enc: A CNN-Enhanced Transformer-Encoder On Top Of Static BERT representations for Document Classification 13 Sep 2022 · 0 repositories · arXiv:2209.06344
-
Robin: A Novel Online Suicidal Text Corpus of Substantial Breadth and Scale 13 Sep 2022 · 0 repositories · arXiv:2209.05707
-
SkIn: Skimming-Intensive Long-Text Classification Using BERT for Medical Corpus 13 Sep 2022 · 0 repositories · arXiv:2209.05741
-
A new hazard event classification model via deep learning and multifractal 12 Sep 2022 · 0 repositories · arXiv:2209.05263
-
DECK: Behavioral Tests to Improve Interpretability and Generalizability of BERT Models Detecting Depression from Text 12 Sep 2022 · 0 repositories · arXiv:2209.05286
-
Chain of Explanation: New Prompting Method to Generate Higher Quality Natural Language Explanation for Implicit Hate Speech 11 Sep 2022 · 0 repositories · arXiv:2209.04889
-
Probing for Understanding of English Verb Classes and Alternations in Large Pre-trained Language Models 11 Sep 2022 · 0 repositories · arXiv:2209.04811
-
Simple and Effective Gradient-Based Tuning of Sequence-to-Sequence Models 10 Sep 2022 · 0 repositories · arXiv:2209.04683
-
Yes, DLGM! A novel hierarchical model for hazard classification 10 Sep 2022 · 0 repositories · arXiv:2209.04576
-
EchoCoTr: Estimation of the Left Ventricular Ejection Fraction from Spatiotemporal Echocardiography 9 Sep 2022 · 1 repository · arXiv:2209.04242
-
Trigger Warnings: Bootstrapping a Violence Detector for FanFiction 9 Sep 2022 · 0 repositories · arXiv:2209.04409
-
CLaCLab at SocialDisNER: Using Medical Gazetteers for Named-Entity Recognition of Disease Mentions in Spanish Tweets 8 Sep 2022 · 1 repository · arXiv:2209.03528
-
IDIAPers @ Causal News Corpus 2022: Extracting Cause-Effect-Signal Triplets via Pre-trained Autoregressive Language Model 8 Sep 2022 · 1 repository · arXiv:2209.03891
-
5q032e@SMM4H'22: Transformer-based classification of premise in tweets related to COVID-19 8 Sep 2022 · 0 repositories · arXiv:2209.03851
-
Why So Toxic? Measuring and Triggering Toxic Behavior in Open-Domain Chatbots 7 Sep 2022 · 0 repositories · arXiv:2209.03463
-
Multilingual Bidirectional Unsupervised Translation Through Multilingual Finetuning and Back-Translation 6 Sep 2022 · 1 repository · arXiv:2209.02821
-
ChemBERTa-2: Towards Chemical Foundation Models 5 Sep 2022 · 2 repositories · arXiv:2209.01712
-
Distilling the Knowledge of BERT for CTC-based ASR 5 Sep 2022 · 0 repositories · arXiv:2209.02030
-
Evaluating the Susceptibility of Pre-Trained Language Models via Handcrafted Adversarial Examples 5 Sep 2022 · 0 repositories · arXiv:2209.02128
-
Do Large Language Models know what humans know? 4 Sep 2022 · 1 repository · arXiv:2209.01515
-
Every picture tells a story: Image-grounded controllable stylistic story generation 4 Sep 2022 · 0 repositories · arXiv:2209.01638
-
Generalization in Neural Networks: A Broad Survey 4 Sep 2022 · 0 repositories · arXiv:2209.01610
-
Elaboration-Generating Commonsense Question Answering at Scale 2 Sep 2022 · 1 repository · arXiv:2209.01232
-
FOLIO: Natural Language Reasoning with First-Order Logic 2 Sep 2022 · 1 repository · arXiv:2209.00840
-
GReS: Graphical Cross-domain Recommendation for Supply Chain Platform 2 Sep 2022 · 0 repositories · arXiv:2209.01031
-
Enhancing Semantic Understanding with Self-supervised Methods for Abstractive Dialogue Summarization 1 Sep 2022 · 0 repositories · arXiv:2209.00278
-
Isotropic Representation Can Improve Dense Retrieval 1 Sep 2022 · 1 repository · arXiv:2209.00218
-
Distilling Multi-Scale Knowledge for Event Temporal Relation Extraction 1 Sep 2022 · 0 repositories · arXiv:2209.00568
-
Negation detection in Dutch clinical texts: an evaluation of rule-based and machine learning methods 1 Sep 2022 · 1 repository · arXiv:2209.00470
-
Few-Shot Learning for Clinical Natural Language Processing Using Siamese Neural Networks 31 Aug 2022 · 0 repositories · arXiv:2208.14923
-
Large-scale Multi-granular Concept Extraction Based on Machine Reading Comprehension 30 Aug 2022 · 1 repository · arXiv:2208.14139
-
No means ‘No’; a non-im-proper modeling approach, with embedded speculative context 30 Aug 2022 · 0 repositories
-
SwiftPruner: Reinforced Evolutionary Pruning for Efficient Ad Relevance 30 Aug 2022 · 0 repositories · arXiv:2209.00625
-
Transformers with Learnable Activation Functions 30 Aug 2022 · 2 repositories · arXiv:2208.14111Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Multi-dimensional Racism Classification during COVID-19: Stigmatization, Offensiveness, Blame, and Exclusion 29 Aug 2022 · 0 repositories · arXiv:2208.13318
-
MDIA: A Benchmark for Multilingual Dialogue Generation in 46 Languages 27 Aug 2022 · 1 repository · arXiv:2208.13078
-
AutoQGS: Auto-Prompt for Low-Resource Knowledge-based Question Generation from SPARQL 26 Aug 2022 · 1 repository · arXiv:2208.12461
-
Building the Intent Landscape of Real-World Conversational Corpora with Extractive Question-Answering Transformers 26 Aug 2022 · 0 repositories · arXiv:2208.12886
-
Task-specific Pre-training and Prompt Decomposition for Knowledge Graph Population with Language Models 26 Aug 2022 · 1 repository · arXiv:2208.12539
-
On Reality and the Limits of Language Data: Aligning LLMs with Human Norms 25 Aug 2022 · 0 repositories · arXiv:2208.11981
-
Training a T5 Using Lab-sized Resources 25 Aug 2022 · 0 repositories · arXiv:2208.12097
-
Addressing Token Uniformity in Transformers via Singular Value Transformation 24 Aug 2022 · 1 repository · arXiv:2208.11790Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Diverse Title Generation for Stack Overflow Posts with Multiple Sampling Enhanced Transformer 24 Aug 2022 · 1 repository · arXiv:2208.11523
-
Evaluate Confidence Instead of Perplexity for Zero-shot Commonsense Reasoning 23 Aug 2022 · 0 repositories · arXiv:2208.11007
-
Prompting as Probing: Using Language Models for Knowledge Base Construction 23 Aug 2022 · 1 repository · arXiv:2208.11057Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
A Syntax Aware BERT for Identifying Well-Formed Queries in a Curriculum Framework 21 Aug 2022 · 0 repositories · arXiv:2208.09912
-
CMSBERT-CLR: Context-driven Modality Shifting BERT with Contrastive Learning for linguistic, visual, acoustic Representations 21 Aug 2022 · 0 repositories · arXiv:2209.07424
-
BSpell: A CNN-Blended BERT Based Bangla Spell Checker 20 Aug 2022 · 1 repository · arXiv:2208.09709
-
Combining Compressions for Multiplicative Size Scaling on Natural Language Tasks 20 Aug 2022 · 0 repositories · arXiv:2208.09684
-
Pretrained Language Encoders are Natural Tagging Frameworks for Aspect Sentiment Triplet Extraction 20 Aug 2022 · 0 repositories · arXiv:2208.09617
-
SPOT: Knowledge-Enhanced Language Representations for Information Extraction 20 Aug 2022 · 0 repositories · arXiv:2208.09625
-
Graph-Augmented Cyclic Learning Framework for Similarity Estimation of Medical Clinical Notes 19 Aug 2022 · 0 repositories · arXiv:2208.09437
-
UniCausal: Unified Benchmark and Repository for Causal Text Mining 19 Aug 2022 · 1 repository · arXiv:2208.09163
-
MulZDG: Multilingual Code-Switching Framework for Zero-shot Dialogue Generation 18 Aug 2022 · 1 repository · arXiv:2208.08629