Methods › General › Regularization › Attention Dropout › Papers, page 69
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 69 of 109: papers 6,801 to 6,900 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Deep learning based Chinese text sentiment mining and stock market correlation research 10 May 2022 · 0 repositories · arXiv:2205.04743
-
Problems with Cosine as a Measure of Embedding Similarity for High Frequency Words 10 May 2022 · 2 repositories · arXiv:2205.05092
-
Ratatouille: A tool for Novel Recipe Generation 10 May 2022 · 0 repositories · arXiv:2206.08267
-
Reducing Activation Recomputation in Large Transformer Models 10 May 2022 · 4 repositories · arXiv:2205.05198Syntology official: harvested, nothing ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
UL2: Unifying Language Learning Paradigms 10 May 2022 · 2 repositories · arXiv:2205.05131Syntology community repositories only · 15 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 16 harvested samples)
-
Automated Evaluation for Student Argumentative Writing: A Survey 9 May 2022 · 0 repositories · arXiv:2205.04083
-
Long Document Re-ranking with Modular Re-ranker 9 May 2022 · 1 repository · arXiv:2205.04275
-
Multi-segment preserving sampling for deep manifold sampler 9 May 2022 · 0 repositories · arXiv:2205.04259
-
Research on the correlation between text emotion mining and stock market based on deep learning 9 May 2022 · 0 repositories · arXiv:2205.06675
-
On the Use of BERT for Automated Essay Scoring: Joint Learning of Multi-Scale Essay Representation 8 May 2022 · 1 repository · arXiv:2205.03835
-
AKI-BERT: a Pre-trained Clinical Language Model for Early Prediction of Acute Kidney Injury 7 May 2022 · 1 repository · arXiv:2205.03695
-
EmotionFlow: Capture the Dialogue Level Emotion Transitions 7 May 2022 · 1 repository
-
Improving Downstream Task Performance by Treating Numbers as Entities 7 May 2022 · 0 repositories · arXiv:2205.03559
-
Vector Representations of Idioms in Conversational Systems 7 May 2022 · 0 repositories · arXiv:2205.03666
-
A Data Cartography based MixUp for Pre-trained Language Models 6 May 2022 · 1 repository · arXiv:2205.03403
-
Stock Price Prediction Based on Natural Language Processing 6 May 2022 · 2 repositories
-
The Unreliability of Explanations in Few-shot Prompting for Textual Reasoning 6 May 2022 · 1 repository · arXiv:2205.03401
-
When a sentence does not introduce a discourse entity, Transformer-based models still sometimes refer to it 6 May 2022 · 1 repository · arXiv:2205.03472
-
BORT: Back and Denoising Reconstruction for End-to-End Task-Oriented Dialog 5 May 2022 · 1 repository · arXiv:2205.02471
-
Exploiting Global and Local Hierarchies for Hierarchical Text Classification 5 May 2022 · 1 repository · arXiv:2205.02613
-
Hyperbolic Relevance Matching for Neural Keyphrase Extraction 4 May 2022 · 1 repository · arXiv:2205.02047
-
Provably Confidential Language Modelling 4 May 2022 · 1 repository · arXiv:2205.01863Syntology official: no sample here; runs from other or unrecorded repositories · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Using virtual edges to extract keywords from texts modeled as complex networks 4 May 2022 · 0 repositories · arXiv:2205.02172
-
Contrastive Learning for Prompt-Based Few-Shot Language Learners 3 May 2022 · 1 repository · arXiv:2205.01308Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples) · 6 pointer-only (licence)
-
Efficient Fine-Tuning of BERT Models on the Edge 3 May 2022 · 0 repositories · arXiv:2205.01541
-
Explain and Conquer: Personalised Text-based Reviews to Achieve Transparency 3 May 2022 · 0 repositories · arXiv:2205.01759
-
Finding patterns in Knowledge Attribution for Transformers 3 May 2022 · 0 repositories · arXiv:2205.01366
-
Mixed-effects transformers for hierarchical adaptation 3 May 2022 · 1 repository · arXiv:2205.01749
-
Predicting Issue Types with seBERT 3 May 2022 · 1 repository · arXiv:2205.01335
-
SemAttack: Natural Textual Attacks via Different Semantic Spaces 3 May 2022 · 1 repository · arXiv:2205.01287Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
BERTops: Studying BERT Representations under a Topological Lens 2 May 2022 · 1 repository · arXiv:2205.00953
-
Entity-aware Transformers for Entity Search 2 May 2022 · 1 repository · arXiv:2205.00820
-
Improving Students' Academic Performance with AI and Semantic Technologies 2 May 2022 · 1 repository · arXiv:2206.03213
-
OPT: Open Pre-trained Transformer Language Models 2 May 2022 · 11 repositories · arXiv:2205.01068Syntology official (archive's flag): 9 ran · 14 ran (of which 2 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 10 unverified (of 24 harvested samples) · 17 pointer-only (licence)
-
Detecting COVID-19 Conspiracy Theories with Transformers and TF-IDF 1 May 2022 · 0 repositories · arXiv:2205.00377
-
StorSeismic: A new paradigm in deep learning for seismic processing 30 Apr 2022 · 1 repository · arXiv:2205.00222
-
ExaASC: A General Target-Based Stance Detection Corpus in Arabic Language 29 Apr 2022 · 1 repository · arXiv:2204.13979
-
Training Language Models with Language Feedback 29 Apr 2022 · 0 repositories · arXiv:2204.14146
-
QRelScore: Better Evaluating Generated Questions with Deeper Understanding of Context-aware Relevance 29 Apr 2022 · 0 repositories · arXiv:2204.13921
-
Two New Datasets for Italian-Language Abstractive Text Summarization 29 Apr 2022 · 1 repository
-
HiNER: A Large Hindi Named Entity Recognition Dataset 28 Apr 2022 · 1 repository · arXiv:2204.13743
-
Inferring Implicit Relations in Complex Questions with Language Models 28 Apr 2022 · 1 repository · arXiv:2204.13778
-
On the Effect of Pretraining Corpora on In-context Learning by a Large-scale Language Model 28 Apr 2022 · 0 repositories · arXiv:2204.13509
-
RobBERTje: a Distilled Dutch BERT Model 28 Apr 2022 · 0 repositories · arXiv:2204.13511
-
Tailor: A Prompt-Based Approach to Attribute-Based Controlled Text Generation 28 Apr 2022 · 0 repositories · arXiv:2204.13362
-
An End-to-End Dialogue Summarization System for Sales Calls 27 Apr 2022 · 0 repositories · arXiv:2204.12951
-
Better Query Graph Selection for Knowledge Base Question Answering 27 Apr 2022 · 0 repositories · arXiv:2204.12662
-
Modern Baselines for SPARQL Semantic Parsing 27 Apr 2022 · 1 repository · arXiv:2204.12793
-
RigoBERTa: A State-of-the-Art Language Model For Spanish 27 Apr 2022 · 0 repositories · arXiv:2205.10233
-
SkillSpan: Hard and Soft Skill Extraction from English Job Postings 27 Apr 2022 · 1 repository · arXiv:2204.12811
-
BiTimeBERT: Extending Pre-Trained Language Representations with Bi-Temporal Information 27 Apr 2022 · 0 repositories · arXiv:2204.13032
-
MILES: Visual BERT Pre-training with Injected Language Semantics for Video-text Retrieval 26 Apr 2022 · 1 repository · arXiv:2204.12408
-
PLOD: An Abbreviation Detection Dataset for Scientific Documents 26 Apr 2022 · 1 repository · arXiv:2204.12061
-
Pretraining Chinese BERT for Detecting Word Insertion and Deletion Errors 26 Apr 2022 · 0 repositories · arXiv:2204.12052
-
Evaluating Interpolation and Extrapolation Performance of Neural Retrieval Models 25 Apr 2022 · 1 repository · arXiv:2204.11447
-
Groupwise Query Performance Prediction with BERT 25 Apr 2022 · 1 repository · arXiv:2204.11489
-
Faster Learned Sparse Retrieval with Guided Traversal 24 Apr 2022 · 1 repository · arXiv:2204.11314
-
Learning to Win Lottery Tickets in BERT Transfer via Task-agnostic Mask Training 24 Apr 2022 · 1 repository · arXiv:2204.11218
-
Fine-Tuning BERT Models to Classify Misinformation on Garlic and COVID-19 on Twitter 22 Apr 2022 · 0 repositories
-
Taygete at SemEval-2022 Task 4: RoBERTa based models for detecting Patronising and Condescending Language 22 Apr 2022 · 0 repositories · arXiv:2204.10519
-
WaBERT: A Low-resource End-to-end Model for Spoken Language Understanding and Speech-to-BERT Alignment 22 Apr 2022 · 0 repositories · arXiv:2204.10461
-
Measuring artificial intelligence: a systematic assessment and implications for governance 21 Apr 2022 · 0 repositories · arXiv:2204.10304
-
Is Neural Topic Modelling Better than Clustering? An Empirical Study on Clustering with Contextual Embeddings for Topics 21 Apr 2022 · 1 repository · arXiv:2204.09874Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Is BERT Robust to Label Noise? A Study on Learning with Noisy Labels in Text Classification 20 Apr 2022 · 1 repository · arXiv:2204.09371Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Towards Arabic Sentence Simplification via Classification and Generative Approaches 20 Apr 2022 · 0 repositories · arXiv:2204.09292
-
ALBETO and DistilBETO: Lightweight Spanish Language Models 19 Apr 2022 · 2 repositories · arXiv:2204.09145
-
CodexDB: Generating Code for Processing SQL Queries using GPT-3 Codex 19 Apr 2022 · 0 repositories · arXiv:2204.08941
-
DecBERT: Enhancing the Language Understanding of BERT with Causal Attention Masks 19 Apr 2022 · 0 repositories · arXiv:2204.08688
-
Impact of Tokenization on Language Models: An Analysis for Turkish 19 Apr 2022 · 0 repositories · arXiv:2204.08832
-
LitMC-BERT: transformer-based multi-label classification of biomedical literature with an application on COVID-19 literature curation 19 Apr 2022 · 1 repository · arXiv:2204.08649
-
Mono vs Multilingual BERT for Hate Speech Detection and Text Classification: A Case Study in Marathi 19 Apr 2022 · 0 repositories · arXiv:2204.08669
-
Multimodal Hate Speech Detection from Bengali Memes and Texts 19 Apr 2022 · 1 repository · arXiv:2204.10196
-
Probing for the Usage of Grammatical Number 19 Apr 2022 · 0 repositories · arXiv:2204.08831
-
Ingredient Extraction from Text in the Recipe Domain 18 Apr 2022 · 1 repository · arXiv:2204.08137
-
L3Cube-HingCorpus and HingBERT: A Code Mixed Hindi-English Dataset and BERT Language Models 18 Apr 2022 · 1 repository · arXiv:2204.08398
-
MASSIVE: A 1M-Example Multilingual Natural Language Understanding Dataset with 51 Typologically-Diverse Languages 18 Apr 2022 · 6 repositories · arXiv:2204.08582Syntology official (archive's flag): 1 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
UTNLP at SemEval-2022 Task 6: A Comparative Analysis of Sarcasm Detection Using Generative-based and Mutation-based Data Augmentation 18 Apr 2022 · 2 repositories · arXiv:2204.08198
-
Zero-shot Entity and Tweet Characterization with Designed Conditional Prompts and Contexts 18 Apr 2022 · 0 repositories · arXiv:2204.08405
-
Knowledgeable Salient Span Mask for Enhancing Language Models as Knowledge Base 17 Apr 2022 · 0 repositories · arXiv:2204.07994
-
Pathologies of Pre-trained Language Models in Few-shot Fine-tuning 17 Apr 2022 · 0 repositories · arXiv:2204.08039
-
Probing Script Knowledge from Pre-Trained Models 16 Apr 2022 · 0 repositories · arXiv:2204.10176
-
SimpleBERT: A Pre-trained Model That Learns to Generate Simple Words 16 Apr 2022 · 0 repositories · arXiv:2204.07779
-
WordAlchemy: A transformer-based Reverse Dictionary 16 Apr 2022 · 0 repositories · arXiv:2204.10181
-
Improving Pre-trained Language Models with Syntactic Dependency Prediction Task for Chinese Semantic Error Recognition 15 Apr 2022 · 0 repositories · arXiv:2204.07464
-
mGPT: Few-Shot Learners Go Multilingual 15 Apr 2022 · 1 repository · arXiv:2204.07580
-
ML_LTU at SemEval-2022 Task 4: T5 Towards Identifying Patronizing and Condescending Language 15 Apr 2022 · 0 repositories · arXiv:2204.07432
-
Polling Latent Opinions: A Method for Computational Sociolinguistics Using Transformer Language Models 15 Apr 2022 · 1 repository · arXiv:2204.07483
-
Resource-Aware Distributed Submodular Maximization: A Paradigm for Multi-Robot Decision-Making 15 Apr 2022 · 0 repositories · arXiv:2204.07520
-
Analysing similarities between legal court documents using natural language processing approaches based on Transformers 14 Apr 2022 · 0 repositories · arXiv:2204.07182
-
CalBERT - Code-mixed Adaptive Language representations using BERT 14 Apr 2022 · 1 repository
-
Does BERT really agree ? Fine-grained Analysis of Lexical Dependence on a Syntactic Task 14 Apr 2022 · 0 repositories · arXiv:2204.06889
-
GPT-NeoX-20B: An Open-Source Autoregressive Language Model 14 Apr 2022 · 11 repositories · arXiv:2204.06745Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Latent Aspect Detection from Online Unsolicited Customer Reviews 14 Apr 2022 · 1 repository · arXiv:2204.06964
-
Multi-label topic classification for COVID-19 literature with Bioformer 14 Apr 2022 · 0 repositories · arXiv:2204.06758
-
Rows from Many Sources: Enriching row completions from Wikidata with a pre-trained Language Model 14 Apr 2022 · 0 repositories · arXiv:2204.07014
-
ViTOL: Vision Transformer for Weakly Supervised Object Localization 14 Apr 2022 · 1 repository · arXiv:2204.06772
-
HuBERT-EE: Early Exiting HuBERT for Efficient Speech Recognition 13 Apr 2022 · 1 repository · arXiv:2204.06328
-
IIITDWD-ShankarB@ Dravidian-CodeMixi-HASOC2021: mBERT based model for identification of offensive content in south Indian languages 13 Apr 2022 · 0 repositories · arXiv:2204.10195
-
METRO: Efficient Denoising Pretraining of Large Scale Autoencoding Language Models with Model Generated Signals 13 Apr 2022 · 0 repositories · arXiv:2204.06644
-
Probing for Constituency Structure in Neural Language Models 13 Apr 2022 · 1 repository · arXiv:2204.06201Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)