Methods › General › Regularization › Attention Dropout › Papers, page 66
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 66 of 109: papers 6,501 to 6,600 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Using Large Language Models to Simulate Multiple Humans and Replicate Human Subject Studies 18 Aug 2022 · 2 repositories · arXiv:2208.10264Syntology community repositories only · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 4 harvested samples)
-
VAuLT: Augmenting the Vision-and-Language Transformer for Sentiment Classification on Social Media 18 Aug 2022 · 1 repository · arXiv:2208.09021Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
EmoMent: An Emotion Annotated Mental Health Corpus from two South Asian Countries 17 Aug 2022 · 0 repositories · arXiv:2208.08486
-
Neural Embeddings for Text 17 Aug 2022 · 1 repository · arXiv:2208.08386
-
Summarizing Patients Problems from Hospital Progress Notes Using Pre-trained Sequence-to-Sequence Models 17 Aug 2022 · 0 repositories · arXiv:2208.08408
-
Transformer Encoder for Social Science 17 Aug 2022 · 1 repository · arXiv:2208.08005
-
Continuous Active Learning Using Pretrained Transformers 15 Aug 2022 · 0 repositories · arXiv:2208.06955
-
MoCapAct: A Multi-Task Dataset for Simulated Humanoid Control 15 Aug 2022 · 1 repository · arXiv:2208.07363Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Targeted Honeyword Generation with Language Models 15 Aug 2022 · 0 repositories · arXiv:2208.06946
-
Teacher Guided Training: An Efficient Framework for Knowledge Transfer 14 Aug 2022 · 0 repositories · arXiv:2208.06825
-
Text Difficulty Study: Do machines behave the same as humans regarding text difficulty? 14 Aug 2022 · 0 repositories · arXiv:2208.14509
-
Adan: Adaptive Nesterov Momentum Algorithm for Faster Optimizing Deep Models 13 Aug 2022 · 9 repositories · arXiv:2208.06677Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Interpreting BERT-based Text Similarity via Activation and Saliency Maps 13 Aug 2022 · 0 repositories · arXiv:2208.06612
-
Is Your Model Sensitive? SPeDaC: A New Benchmark for Detecting and Classifying Sensitive Personal Data 12 Aug 2022 · 0 repositories · arXiv:2208.06216
-
Pre-training Tasks for User Intent Detection and Embedding Retrieval in E-commerce Search 12 Aug 2022 · 1 repository · arXiv:2208.06150
-
A Model of Anaphoric Ambiguities using Sheaf Theoretic Quantum-like Contextuality and BERT 11 Aug 2022 · 0 repositories · arXiv:2208.05720
-
A Twitter-Driven Deep Learning Mechanism for the Determination of Vehicle Hijacking Spots in Cities 11 Aug 2022 · 0 repositories · arXiv:2208.10280
-
Searching for chromate replacements using natural language processing and machine learning algorithms 11 Aug 2022 · 0 repositories · arXiv:2208.05672
-
A Boring-yet-effective Approach for the Product Ranking Task of the Amazon KDD Cup 2022 9 Aug 2022 · 0 repositories · arXiv:2208.06264
-
A Multimodal Transformer: Fusing Clinical Notes with Structured EHR Data for Interpretable In-Hospital Mortality Prediction 9 Aug 2022 · 0 repositories · arXiv:2208.10240
-
E2EG: End-to-End Node Classification Using Graph Topology and Text-based Node Attributes 9 Aug 2022 · 1 repository · arXiv:2208.04609Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 4 pointer-only (licence)
-
Emotion Detection From Tweets Using a BERT and SVM Ensemble Model 9 Aug 2022 · 1 repository · arXiv:2208.04547
-
Exploring Hate Speech Detection with HateXplain and BERT 9 Aug 2022 · 1 repository · arXiv:2208.04489
-
Debiased Large Language Models Still Associate Muslims with Uniquely Violent Acts 8 Aug 2022 · 0 repositories · arXiv:2208.04417
-
Efficient Fine-Tuning of Compressed Language Models with Learners 3 Aug 2022 · 0 repositories · arXiv:2208.02070
-
A Comparative Study on COVID-19 Fake News Detection Using Different Transformer Based Models 2 Aug 2022 · 0 repositories · arXiv:2208.01355
-
Automatic Classification of Bug Reports Based on Multiple Text Information and Reports' Intention 2 Aug 2022 · 0 repositories · arXiv:2208.01274
-
Debiasing Gender Bias in Information Retrieval Models 2 Aug 2022 · 0 repositories · arXiv:2208.01755
-
giMLPs: Gate with Inhibition Mechanism in MLPs 1 Aug 2022 · 1 repository · arXiv:2208.00929
-
Interacting with next-phrase suggestions: How suggestion systems aid and influence the cognitive processes of writing 1 Aug 2022 · 0 repositories · arXiv:2208.00636
-
What Can Transformers Learn In-Context? A Case Study of Simple Function Classes 1 Aug 2022 · 2 repositories · arXiv:2208.01066
-
Aggretriever: A Simple Approach to Aggregate Textual Representations for Robust Dense Passage Retrieval 31 Jul 2022 · 1 repository · arXiv:2208.00511
-
Neural Knowledge Bank for Pretrained Transformers 31 Jul 2022 · 0 repositories · arXiv:2208.00399
-
A Survey on Masked Autoencoder for Self-supervised Learning in Vision and Beyond 30 Jul 2022 · 0 repositories · arXiv:2208.00173
-
Code Comment Inconsistency Detection with BERT and Longformer 29 Jul 2022 · 1 repository · arXiv:2207.14444
-
Curriculum Learning for Data-Efficient Vision-Language Alignment 29 Jul 2022 · 0 repositories · arXiv:2207.14525
-
SERCNN: Stacked Embedding Recurrent Convolutional Neural Network in Detecting Depression on Twitter 29 Jul 2022 · 0 repositories · arXiv:2207.14535
-
CrAM: A Compression-Aware Minimizer 28 Jul 2022 · 1 repository · arXiv:2207.14200Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples)
-
LAD: Language Models as Data for Zero-Shot Dialog 28 Jul 2022 · 0 repositories · arXiv:2207.14393
-
Large Language Models and the Reverse Turing Test 28 Jul 2022 · 0 repositories · arXiv:2207.14382
-
SDBERT: SparseDistilBERT, a faster and smaller BERT model 28 Jul 2022 · 0 repositories · arXiv:2208.10246
-
Sequence to sequence pretraining for a less-resourced Slovenian language 28 Jul 2022 · 1 repository · arXiv:2207.13988
-
SoundChoice: Grapheme-to-Phoneme Models with Semantic Disambiguation 27 Jul 2022 · 1 repository · arXiv:2207.13703
-
Bundle MCR: Towards Conversational Bundle Recommendation 26 Jul 2022 · 1 repository · arXiv:2207.12628
-
Fine-Tuning BERT for Automatic ADME Semantic Labeling in FDA Drug Labeling to Enhance Product-Specific Guidance Assessment 25 Jul 2022 · 0 repositories · arXiv:2207.12376
-
Is GPT-3 all you need for Visual Question Answering in Cultural Heritage? 25 Jul 2022 · 0 repositories · arXiv:2207.12101
-
A Cognitive Study on Semantic Similarity Analysis of Large Corpora: A Transformer-based Approach 24 Jul 2022 · 0 repositories · arXiv:2207.11716
-
No More Fine-Tuning? An Experimental Evaluation of Prompt Tuning in Code Intelligence 24 Jul 2022 · 1 repository · arXiv:2207.11680Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 7 unverified (of 14 harvested samples) · 2 pointer-only (licence)
-
Better Reasoning Behind Classification Predictions with BERT for Fake News Detection 23 Jul 2022 · 0 repositories · arXiv:2207.11562
-
Context based lemmatizer for Polish language 23 Jul 2022 · 0 repositories · arXiv:2207.11565
-
Zero-Shot Video Captioning with Evolving Pseudo-Tokens 22 Jul 2022 · 1 repository · arXiv:2207.11100
-
BigIssue: A Realistic Bug Localization Benchmark 21 Jul 2022 · 0 repositories · arXiv:2207.10739
-
Efficient model compression with Random Operation Access Specific Tile (ROAST) hashing 21 Jul 2022 · 1 repository · arXiv:2207.10702
-
Locality Guidance for Improving Vision Transformers on Tiny Datasets 20 Jul 2022 · 1 repository · arXiv:2207.10026Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Enhancing Collaborative Filtering Recommender with Prompt-Based Sentiment Analysis 19 Jul 2022 · 1 repository · arXiv:2207.12883
-
PiC: A Phrase-in-Context Dataset for Phrase Understanding and Semantic Search 19 Jul 2022 · 1 repository · arXiv:2207.09068
-
Pre-trained language models with domain knowledge for biomedical extractive summarization 19 Jul 2022 · 1 repository
-
Revealing Secrets From Pre-trained Models 19 Jul 2022 · 0 repositories · arXiv:2207.09539
-
Selection Bias Induced Spurious Correlations in Large Language Models 18 Jul 2022 · 1 repository · arXiv:2207.08982
-
Word Play for Playing Othello (Reverses) 18 Jul 2022 · 0 repositories · arXiv:2207.08766
-
Aspect-specific Context Modeling for Aspect-based Sentiment Analysis 17 Jul 2022 · 1 repository · arXiv:2207.08099
-
Can large language models reason about medical questions? 17 Jul 2022 · 1 repository · arXiv:2207.08143Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples)
-
Effectiveness of French Language Models on Abstractive Dialogue Summarization Task 17 Jul 2022 · 0 repositories · arXiv:2207.08305
-
ELECTRA is a Zero-Shot Learner, Too 17 Jul 2022 · 1 repository · arXiv:2207.08141Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Representation Learning of Image Schema 17 Jul 2022 · 0 repositories · arXiv:2207.08256
-
Robust Action Governor for Uncertain Piecewise Affine Systems with Non-convex Constraints and Safe Reinforcement Learning 17 Jul 2022 · 0 repositories · arXiv:2207.08240
-
A Context-Sensitive Word Embedding Approach for The Detection of Troll Tweets 17 Jul 2022 · 0 repositories · arXiv:2207.08230
-
Lightweight Vision Transformer with Cross Feature Attention 15 Jul 2022 · 0 repositories · arXiv:2207.07268
-
POET: Training Neural Networks on Tiny Devices with Integrated Rematerialization and Paging 15 Jul 2022 · 1 repository · arXiv:2207.07697
-
Position Prediction as an Effective Pretraining Strategy 15 Jul 2022 · 1 repository · arXiv:2207.07611
-
Z-Index at CheckThat! Lab 2022: Check-Worthiness Identification on Tweet Text 15 Jul 2022 · 0 repositories · arXiv:2207.07308
-
Combing for Credentials: Active Pattern Extraction from Smart Reply 14 Jul 2022 · 0 repositories · arXiv:2207.10802
-
Bootstrapped Masked Autoencoders for Vision BERT Pretraining 14 Jul 2022 · 1 repository · arXiv:2207.07116Syntology official (archive's flag): 6 ran · 6 ran (of which 6 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 6 samples that ran constructed an object rather than computing a result (of 7 harvested samples) · 7 pointer-only (licence)
-
Language Modelling with Pixels 14 Jul 2022 · 1 repository · arXiv:2207.06991Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Multilinguals at SemEval-2022 Task 11: Complex NER in Semantically Ambiguous Settings for Low Resource Languages 14 Jul 2022 · 1 repository · arXiv:2207.06882
-
A Transfer Learning Based Model for Text Readability Assessment in German 13 Jul 2022 · 0 repositories · arXiv:2207.06265
-
DocPrompting: Generating Code by Retrieving the Docs 13 Jul 2022 · 2 repositories · arXiv:2207.05987Syntology official (archive's flag): 4 ran · 4 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 9 harvested samples)
-
DynaST: Dynamic Sparse Transformer for Exemplar-Guided Image Generation 13 Jul 2022 · 1 repository · arXiv:2207.06124Syntology official (archive's flag): 5 ran · 5 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples)
-
Exploiting Word Semantics to Enrich Character Representations of Chinese Pre-trained Models 13 Jul 2022 · 1 repository · arXiv:2207.05928
-
Re2G: Retrieve, Rerank, Generate 13 Jul 2022 · 1 repository · arXiv:2207.06300Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
How Do Multilingual Encoders Learn Cross-lingual Representation? 12 Jul 2022 · 0 repositories · arXiv:2207.05737
-
Using Paraphrases to Study Properties of Contextual Embeddings 12 Jul 2022 · 0 repositories · arXiv:2207.05553
-
Learning Large-scale Universal User Representation with Sparse Mixture of Experts 11 Jul 2022 · 0 repositories · arXiv:2207.04648
-
Multi-level Fusion of Wav2vec 2.0 and BERT for Multimodal Emotion Recognition 11 Jul 2022 · 1 repository · arXiv:2207.04697
-
Overview of the Shared Task on Fake News Detection in Urdu at FIRE 2021 11 Jul 2022 · 0 repositories · arXiv:2207.05133
-
SparseTIR: Composable Abstractions for Sparse Compilation in Deep Learning 11 Jul 2022 · 2 repositories · arXiv:2207.04606Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 2 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Multilingual Persuasion Detection: Video Games as an Invaluable Data Source for NLP 10 Jul 2022 · 1 repository · arXiv:2207.04453
-
Few-shot training LLMs for project-specific code-summarization 9 Jul 2022 · 0 repositories · arXiv:2207.04237
-
Computationally Identifying Funneling and Focusing Questions in Classroom Discourse 8 Jul 2022 · 1 repository · arXiv:2208.04715
-
Deep Visual-Linguistic Fusion Network Considering Cross-Modal Inconsistency for Rumor Detection 8 Jul 2022 · 3 repositories
-
Hidden Schema Networks 8 Jul 2022 · 0 repositories · arXiv:2207.03777
-
A Large Scale Search Dataset for Unbiased Learning to Rank 7 Jul 2022 · 1 repository · arXiv:2207.03051
-
Active Learning and Multi-label Classification for Ellipsis and Coreference Detection in Conversational Question-Answering 7 Jul 2022 · 0 repositories · arXiv:2207.03145
-
AsNER -- Annotated Dataset and Baseline for Assamese Named Entity recognition 7 Jul 2022 · 0 repositories · arXiv:2207.03422
-
Neural Language Models are not Born Equal to Fit Brain Data, but Training Helps 7 Jul 2022 · 0 repositories · arXiv:2207.03380
-
Sensitivity Analysis on Transferred Neural Architectures of BERT and GPT-2 for Financial Sentiment Analysis 7 Jul 2022 · 0 repositories · arXiv:2207.03037
-
Ask Me What You Need: Product Retrieval using Knowledge from GPT-3 6 Jul 2022 · 0 repositories · arXiv:2207.02516
-
Aspect-Based Sentiment Analysis using Local Context Focus Mechanism with DeBERTa 6 Jul 2022 · 0 repositories · arXiv:2207.02424
-
Learning to Diversify for Product Question Generation 6 Jul 2022 · 0 repositories · arXiv:2207.02534
-
SimLM: Pre-training with Representation Bottleneck for Dense Passage Retrieval 6 Jul 2022 · 1 repository · arXiv:2207.02578