Methods › General › Regularization › Attention Dropout › Papers, page 73
Attention Dropout
Papers archive 2025-07-28
archive papers tagged: 10,892 · with a code link: 4,634 · where Syntology ran a sample: 1,270 (1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,270 of 10,892 tagged: 1,043 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 73 of 109: papers 7,201 to 7,300 of 10,892, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Exploring the Low-Resource Transfer-Learning with mT5 model 16 Jan 2022 · 0 repositories
-
Feasibility of BERT Embeddings For Domain-Specific Knowledge Mining 16 Jan 2022 · 0 repositories
-
FedNLP: Benchmarking Federated Learning Methods for Natural Language Processing Tasks 16 Jan 2022 · 0 repositories
-
Few-Shot Semantic Parsing with Language Models Trained On Code 16 Jan 2022 · 0 repositories
-
Global Entity Disambiguation with BERT 16 Jan 2022 · 0 repositories
-
Hierarchical Transformers Are More Efficient Language Models 16 Jan 2022 · 0 repositories
-
Identifying the Source of Vulnerability in Fragile Interpretations: A Case Study in Neural Text Classification 16 Jan 2022 · 0 repositories
-
IMPLI: Investigatng NLI Models' Performance on Figurative Language 16 Jan 2022 · 0 repositories
-
Improving Contextual Representation with Gloss Regularized Pre-training 16 Jan 2022 · 0 repositories
-
Investigating and Explaining Feature and Representation Learning in Translationese Classification 16 Jan 2022 · 0 repositories
-
Investigating the saliency of sentiment expressions in aspect-based sentiment analysis 16 Jan 2022 · 0 repositories
-
Jointly Reinforced User Simulator and Task-oriented Dialog System with Simplified Generative Architecture 16 Jan 2022 · 0 repositories
-
Language Models for Code-switch Detection of te reo Māori and English in a Low-resource Setting 16 Jan 2022 · 0 repositories
-
Magic Pyramid: Accelerating Inference with Early Exiting and Token Pruning 16 Jan 2022 · 0 repositories
-
Measuring Word-Context Biases in Lexical Semantic Datasets 16 Jan 2022 · 0 repositories
-
Memory-assisted prompt editing to improve GPT-3 after deployment 16 Jan 2022 · 1 repository · arXiv:2201.06009Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Minimally-Supervised Relation Induction from Pre-trained Language Model 16 Jan 2022 · 0 repositories
-
Modeling Multi-Granularity Hierarchical Features for Relation Extraction 16 Jan 2022 · 0 repositories
-
Multi-Stage Pre-Training for Math-Understanding: μ²(AL)BERT 16 Jan 2022 · 0 repositories
-
Natural Language Deduction through Search over Statement Compositions 16 Jan 2022 · 0 repositories · arXiv:2201.06028
-
PCEE-BERT: Accelerating BERT Inference via Patient and Confident Early Exiting 16 Jan 2022 · 0 repositories
-
Penguins Don’t Fly: Reasoning about Generics through Instantiations and Exceptions 16 Jan 2022 · 0 repositories
-
Polling Latent Opinions: A Method for Computational Sociolinguistics Using Transformer Language Models 16 Jan 2022 · 0 repositories
-
Probing The Linguistic Capacity of Pre-Trained Vision-Language Models 16 Jan 2022 · 0 repositories
-
Progressive Class Semantic Matching for Semi-supervised Text Classification 16 Jan 2022 · 0 repositories
-
Provably Confidential Language Modelling 16 Jan 2022 · 0 repositories
-
Re2G: Retrieve, Rerank, Generate 16 Jan 2022 · 0 repositories
-
Reframing Human-AI Collaboration for Generating Free-Text Explanations 16 Jan 2022 · 0 repositories
-
Representation Learning for Conversational Data using Discourse Mutual Information Maximization 16 Jan 2022 · 0 repositories
-
Revisiting Additive Compositionality: AND, OR, and NOT Operations with Word Embeddings 16 Jan 2022 · 0 repositories
-
Robin: A Novel Online Suicidal Text Corpus of Substantial Breadth and Scale 16 Jan 2022 · 0 repositories
-
Roof-BERT: Divide Understanding Labour and Join in Work 16 Jan 2022 · 0 repositories
-
SemAttack: Natural Textual Attacks via Different Semantic Spaces 16 Jan 2022 · 0 repositories
-
Seq-GAN-BERT:Sequence Generative Adversarial Learning for Low-resource Name Entity Recognition 16 Jan 2022 · 0 repositories
-
Simple Local Attentions Remain Competitive for Long-Context Tasks 16 Jan 2022 · 0 repositories
-
TaCL: Improving BERT Pre-training with Token-aware Contrastive Learning 16 Jan 2022 · 0 repositories
-
Tapping BERT for Preposition Sense Disambiguation 16 Jan 2022 · 0 repositories
-
TEMPLATE: TempRel Classification Model Trained with Embedded Temporal Relation Knowledge 16 Jan 2022 · 0 repositories
-
That is a good looking car !: Visual Aspect based Sentiment Controlled Personalized Response Generation 16 Jan 2022 · 0 repositories
-
Tree Knowledge Distillation for Compressing Transformer-Based Language Models 16 Jan 2022 · 0 repositories
-
Uncovering Surprising Event Boundaries in Narratives 16 Jan 2022 · 0 repositories
-
Understand before Answer: Improve Temporal Reading Comprehension via Precise Question Understanding 16 Jan 2022 · 0 repositories
-
UnifiedSKG: Unifying and Multi-Tasking Structured Knowledge Grounding with Text-to-Text Language Models 16 Jan 2022 · 1 repository · arXiv:2201.05966Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
VEE-BERT: Accelerating BERT Inference for Named Entity Recognition via Vote Early Exiting 16 Jan 2022 · 0 repositories
-
WANLI: Worker and AI Collaboration for Natural Language Inference Dataset Creation 16 Jan 2022 · 1 repository · arXiv:2201.05955
-
What do tokens know about their characters and how do they know it? 16 Jan 2022 · 1 repository
-
What Role Does BERT Play in the Neural Machine Translation Encoder? 16 Jan 2022 · 0 repositories
-
When a sentence does not introduce a discourse entity, Transformer-based models still often refer to it 16 Jan 2022 · 0 repositories
-
Why Does Surprisal From Smaller GPT-2 Models Provide Better Fit to Human Reading Times? 16 Jan 2022 · 0 repositories
-
Automatic Correction of Syntactic Dependency Annotation Differences 15 Jan 2022 · 0 repositories · arXiv:2201.05891
-
Automatic Lexical Simplification for Turkish 15 Jan 2022 · 0 repositories · arXiv:2201.05878
-
Machine Learning for Food Review and Recommendation 15 Jan 2022 · 0 repositories · arXiv:2201.10978
-
CommonsenseQA 2.0: Exposing the Limits of AI through Gamification 14 Jan 2022 · 0 repositories · arXiv:2201.05320
-
Polarity and Subjectivity Detection with Multitask Learning and BERT Embedding 14 Jan 2022 · 0 repositories · arXiv:2201.05363
-
Assemble Foundation Models for Automatic Code Summarization 13 Jan 2022 · 1 repository · arXiv:2201.05222
-
Knowledge Graph Augmented Network Towards Multiview Representation Learning for Aspect-based Sentiment Analysis 13 Jan 2022 · 1 repository · arXiv:2201.04831
-
Multi-task Pre-training Language Model for Semantic Network Completion 13 Jan 2022 · 1 repository · arXiv:2201.04843
-
Towards Automated Error Analysis: Learning to Characterize Errors 13 Jan 2022 · 0 repositories · arXiv:2201.05017
-
Diagnosing BERT with Retrieval Heuristics 12 Jan 2022 · 1 repository · arXiv:2201.04458
-
Generative Adversarial Network for Text-to-Face Synthesis and Manipulation with Pretrained BERT Model 12 Jan 2022 · 0 repositories
-
PromptBERT: Improving BERT Sentence Embeddings with Prompts 12 Jan 2022 · 1 repository · arXiv:2201.04337Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified; the one sample that ran constructed an object rather than computing a result (of 5 harvested samples) · 5 pointer-only (licence)
-
A Feature Extraction based Model for Hate Speech Identification 11 Jan 2022 · 0 repositories · arXiv:2201.04227
-
Explaining Predictive Uncertainty by Looking Back at Model Explanations 11 Jan 2022 · 0 repositories · arXiv:2201.03742
-
Quantifying Robustness to Adversarial Word Substitutions 11 Jan 2022 · 0 repositories · arXiv:2201.03829
-
BERT for Sentiment Analysis: Pre-trained and Fine-Tuned Alternatives 10 Jan 2022 · 2 repositories · arXiv:2201.03382
-
Black-Box Tuning for Language-Model-as-a-Service 10 Jan 2022 · 2 repositories · arXiv:2201.03514Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Handwriting recognition and automatic scoring for descriptive answers in Japanese language tests 10 Jan 2022 · 0 repositories · arXiv:2201.03215
-
SCROLLS: Standardized CompaRison Over Long Language Sequences 10 Jan 2022 · 2 repositories · arXiv:2201.03533Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Latency Adjustable Transformer Encoder for Language Understanding 10 Jan 2022 · 0 repositories · arXiv:2201.03327
-
Imagined versus Remembered Stories: Quantifying Differences in Narrative Flow 7 Jan 2022 · 0 repositories · arXiv:2201.02662
-
Flow-Guided Sparse Transformer for Video Deblurring 6 Jan 2022 · 1 repository · arXiv:2201.01893Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Self-Training Vision Language BERTs with a Unified Conditional Model 6 Jan 2022 · 0 repositories · arXiv:2201.02010
-
Formal Analysis of Art: Proxy Learning of Visual Concepts from Style Through Language Models 5 Jan 2022 · 0 repositories · arXiv:2201.01819
-
Learning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction 5 Jan 2022 · 2 repositories · arXiv:2201.02184
-
Comparison of biomedical relationship extraction methods and models for knowledge graph creation 5 Jan 2022 · 0 repositories · arXiv:2201.01647
-
Submix: Practical Private Prediction for Large-Scale Language Models 4 Jan 2022 · 0 repositories · arXiv:2201.00971
-
An Adversarial Benchmark for Fake News Detection Models 3 Jan 2022 · 1 repository · arXiv:2201.00912
-
Which Student is Best? A Comprehensive Knowledge Distillation Exam for Task-Specific BERT Models 3 Jan 2022 · 0 repositories · arXiv:2201.00558
-
On Sensitivity of Deep Learning Based Text Classification Algorithms to Practical Input Perturbations 2 Jan 2022 · 0 repositories · arXiv:2201.00318
-
Continual Stereo Matching of Continuous Driving Scenes With Growing Architecture 1 Jan 2022 · 1 repository
-
Expanding Large Pre-Trained Unimodal Models With Multimodal Information Injection for Image-Text Multimodal Classification 1 Jan 2022 · 0 repositories
-
SpaceEdit: Learning a Unified Editing Space for Open-Domain Image Color Editing 1 Jan 2022 · 0 repositories
-
A Neural Network Solves, Explains, and Generates University Math Problems by Program Synthesis and Few-Shot Learning at Human Level 31 Dec 2021 · 1 repository · arXiv:2112.15594Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Clustering Vietnamese Conversations From Facebook Page To Build Training Dataset For Chatbot 31 Dec 2021 · 1 repository · arXiv:2112.15338
-
Multi-Dimensional Model Compression of Vision Transformer 31 Dec 2021 · 1 repository · arXiv:2201.00043
-
Automatic Mixed-Precision Quantization Search of BERT 30 Dec 2021 · 0 repositories · arXiv:2112.14938
-
EvoMoE: An Evolutional Mixture-of-Experts Training Framework via Dense-To-Sparse Gate 29 Dec 2021 · 2 repositories · arXiv:2112.14397Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Universal Transformer Hawkes Process with Adaptive Recursive Iteration 29 Dec 2021 · 0 repositories · arXiv:2112.14479
-
The University of Texas at Dallas HLTRI's Participation in EPIC-QA: Searching for Entailed Questions Revealing Novel Answer Nuggets 28 Dec 2021 · 0 repositories · arXiv:2112.13946
-
"A Passage to India": Pre-trained Word Embeddings for Indian Languages 27 Dec 2021 · 0 repositories · arXiv:2112.13800
-
Contextual Sentence Analysis for the Sentiment Prediction on Financial Data 27 Dec 2021 · 0 repositories · arXiv:2112.13790
-
Event-based clinical findings extraction from radiology reports with pre-trained language model 27 Dec 2021 · 1 repository · arXiv:2112.13512
-
Mind the Gap: Cross-Lingual Information Retrieval with Hierarchical Knowledge Enhancement 27 Dec 2021 · 0 repositories · arXiv:2112.13510
-
Multi-Image Visual Question Answering 27 Dec 2021 · 1 repository · arXiv:2112.13706
-
Secondary Use of Clinical Problem List Entries for Neural Network-Based Disease Code Assignment 27 Dec 2021 · 0 repositories · arXiv:2112.13756
-
Evaluating Contextual Embeddings and their Extraction Layers for Depression Assessment 27 Dec 2021 · 0 repositories · arXiv:2112.13795
-
An Ensemble of Pre-trained Transformer Models For Imbalanced Multiclass Malware Classification 25 Dec 2021 · 1 repository · arXiv:2112.13236
-
CABACE: Injecting Character Sequence Information and Domain Knowledge for Enhanced Acronym and Long-Form Extraction 25 Dec 2021 · 1 repository · arXiv:2112.13237
-
Deeper Clinical Document Understanding Using Relation Extraction 25 Dec 2021 · 1 repository · arXiv:2112.13259
-
Distilling the Knowledge of Romanian BERTs Using Multiple Teachers 23 Dec 2021 · 1 repository · arXiv:2112.12650