Methods › General › Regularization › Dropout › Papers, page 225
Dropout
Papers archive 2025-07-28
archive papers tagged: 27,472 · with a code link: 12,129 · where Syntology ran a sample: 3,620 (3,044 with a run with no instrument failure, 576 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (3,620 of 27,472 tagged: 3,044 with a run with no instrument failure, 576 where every run was a failure of Syntology's instrument)
Page 225 of 275: papers 22,401 to 22,500 of 27,472, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
On Position Embeddings in BERT 1 Jan 2021 · 0 repositories
-
Parameterization of Hypercomplex Multiplications 1 Jan 2021 · 0 repositories
-
PhraseTransformer: Self-Attention using Local Context for Semantic Parsing 1 Jan 2021 · 1 repository
-
Polyjuice: Generating Counterfactuals for Explaining, Evaluating, and Improving Models 1 Jan 2021 · 1 repository · arXiv:2101.00288
-
Post-Training Weighted Quantization of Neural Networks for Language Models 1 Jan 2021 · 0 repositories
-
Pre-training Text-to-Text Transformers to Write and Reason with Concepts 1 Jan 2021 · 0 repositories
-
Predictive Attention Transformer: Improving Transformer with Attention Map Prediction 1 Jan 2021 · 0 repositories
-
Prefix-Tuning: Optimizing Continuous Prompts for Generation 1 Jan 2021 · 13 repositories · arXiv:2101.00190Syntology community repositories only · 4 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Pretrain Knowledge-Aware Language Models 1 Jan 2021 · 0 repositories
-
Quantile Regularization : Towards Implicit Calibration of Regression Models 1 Jan 2021 · 0 repositories
-
Representation and Bias in Multilingual NLP: Insights from Controlled Experiments on Conditional Language Modeling 1 Jan 2021 · 0 repositories
-
Representational correlates of hierarchical phrase structure in deep language models 1 Jan 2021 · 0 repositories
-
Scene Context-Aware Salient Object Detection 1 Jan 2021 · 1 repository
-
SEDONA: Search for Decoupled Neural Networks toward Greedy Block-wise Learning 1 Jan 2021 · 1 repository
-
Self-Born Wiring for Neural Trees 1 Jan 2021 · 0 repositories
-
Semi-Relaxed Quantization with DropBits: Training Low-Bit Neural Networks via Bitwise Regularization 1 Jan 2021 · 0 repositories
-
Share or Not? Learning to Schedule Language-Specific Capacity for Multilingual Translation 1 Jan 2021 · 0 repositories
-
Single Layers of Attention Suffice to Predict Protein Contacts 1 Jan 2021 · 0 repositories
-
SkillBERT: “Skilling” the BERT to classify skills! 1 Jan 2021 · 0 repositories
-
Speeding up Deep Learning Training by Sharing Weights and Then Unsharing 1 Jan 2021 · 0 repositories
-
STAR: A Structure-Aware Lightweight Transformer for Real-Time Image Enhancement 1 Jan 2021 · 1 repository
-
Stochastic Partial Swap: Enhanced Model Generalization and Interpretability for Fine-Grained Recognition 1 Jan 2021 · 1 repository
-
Subformer: A Parameter Reduced Transformer 1 Jan 2021 · 0 repositories
-
Subformer: Exploring Weight Sharing for Parameter Efficiency in Generative Transformers 1 Jan 2021 · 1 repository · arXiv:2101.00234
-
Syntactic Relevance XLNet Word Embedding Generation in Low-Resource Machine Translation 1 Jan 2021 · 0 repositories
-
Synthesising Realistic Calcium Imaging Data of Neuronal Populations Using GAN 1 Jan 2021 · 1 repository
-
Synthesizer: Rethinking Self-Attention for Transformer Models 1 Jan 2021 · 0 repositories
-
Taking Notes on the Fly Helps Language Pre-Training 1 Jan 2021 · 0 repositories
-
Task-Agnostic and Adaptive-Size BERT Compression 1 Jan 2021 · 0 repositories
-
Towards Practical Second Order Optimization for Deep Learning 1 Jan 2021 · 0 repositories
-
Towards Understanding and Improving Dropout in Game Theory 1 Jan 2021 · 0 repositories
-
Trans-Caps: Transformer Capsule Networks with Self-attention Routing 1 Jan 2021 · 0 repositories
-
Transformer protein language models are unsupervised structure learners 1 Jan 2021 · 0 repositories
-
Transformer-QL: A Step Towards Making Transformer Network Quadratically Large 1 Jan 2021 · 0 repositories
-
Transformers satisfy 1 Jan 2021 · 0 repositories
-
Transforming Recurrent Neural Networks with Attention and Fixed-point Equations 1 Jan 2021 · 0 repositories
-
TRAR: Routing the Attention Spans in Transformer for Visual Question Answering 1 Jan 2021 · 1 repository
-
U-BERT: Pre-training User Representations for Improved Recommendation 1 Jan 2021 · 0 repositories
-
UPDeT: Universal Multi-agent RL via Policy Decoupling with Transformers 1 Jan 2021 · 0 repositories
-
UserBERT: Self-supervised User Representation Learning 1 Jan 2021 · 0 repositories
-
Visual Transformers: Where Do Transformers Really Belong in Vision Models? 1 Jan 2021 · 0 repositories
-
VisualSparta: An Embarrassingly Simple Approach to Large-scale Text-to-Image Search with Weighted Bag-of-words 1 Jan 2021 · 1 repository · arXiv:2101.00265
-
WARP: Word-level Adversarial ReProgramming 1 Jan 2021 · 1 repository · arXiv:2101.00121
-
WB-DETR: Transformer-Based Detector Without Backbone 1 Jan 2021 · 0 repositories
-
A Closer Look at Few-Shot Crosslingual Transfer: The Choice of Shots Matters 31 Dec 2020 · 0 repositories · arXiv:2012.15682
-
A Multi-modal Deep Learning Model for Video Thumbnail Selection 31 Dec 2020 · 0 repositories · arXiv:2101.00073
-
Better Robustness by More Coverage: Adversarial Training with Mixup Augmentation for Robust Fine-tuning 31 Dec 2020 · 1 repository · arXiv:2012.15699
-
BinaryBERT: Pushing the Limit of BERT Quantization 31 Dec 2020 · 1 repository · arXiv:2012.15701Syntology official (archive's flag): 1 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
CoCoLM: COmplex COmmonsense Enhanced Language Model with Discourse Relations 31 Dec 2020 · 1 repository · arXiv:2012.15643
-
Conditional Generation of Temporally-ordered Event Sequences 31 Dec 2020 · 0 repositories · arXiv:2012.15786
-
Directed Beam Search: Plug-and-Play Lexically Constrained Language Generation 31 Dec 2020 · 1 repository · arXiv:2012.15416Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
EarlyBERT: Efficient BERT Training via Early-bird Lottery Tickets 31 Dec 2020 · 1 repository · arXiv:2101.00063Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Estimating Uncertainty in Neural Networks for Cardiac MRI Segmentation: A Benchmark Study 31 Dec 2020 · 0 repositories · arXiv:2012.15772
-
Evidence-based Factual Error Correction 31 Dec 2020 · 3 repositories · arXiv:2012.15788
-
Fully Non-autoregressive Neural Machine Translation: Tricks of the Trade 31 Dec 2020 · 1 repository · arXiv:2012.15833
-
Graph Networks with Spectral Message Passing 31 Dec 2020 · 0 repositories · arXiv:2101.00079
-
KART: Parameterization of Privacy Leakage Scenarios from Pre-trained Language Models 31 Dec 2020 · 1 repository · arXiv:2101.00036
-
Fast WordPiece Tokenization 31 Dec 2020 · 1 repository · arXiv:2012.15524
-
Making Pre-trained Language Models Better Few-shot Learners 31 Dec 2020 · 9 repositories · arXiv:2012.15723Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 9 harvested samples) · 7 pointer-only (licence)
-
MiniLMv2: Multi-Head Self-Attention Relation Distillation for Compressing Pretrained Transformers 31 Dec 2020 · 2 repositories · arXiv:2012.15828
-
Revisiting Robust Neural Machine Translation: A Transformer Case Study 31 Dec 2020 · 0 repositories · arXiv:2012.15710
-
Studying Strategically: Learning to Mask for Closed-book QA 31 Dec 2020 · 0 repositories · arXiv:2012.15856
-
The Pile: An 800GB Dataset of Diverse Text for Language Modeling 31 Dec 2020 · 22 repositories · arXiv:2101.00027Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
TransTrack: Multiple Object Tracking with Transformer 31 Dec 2020 · 2 repositories · arXiv:2012.15460
-
Unified Mandarin TTS Front-end Based on Distilled BERT Model 31 Dec 2020 · 1 repository · arXiv:2012.15404
-
UNKs Everywhere: Adapting Multilingual Language Models to New Scripts 31 Dec 2020 · 2 repositories · arXiv:2012.15562
-
Verb Knowledge Injection for Multilingual Event Processing 31 Dec 2020 · 0 repositories · arXiv:2012.15421
-
XLM-T: Scaling up Multilingual Machine Translation with Pretrained Cross-lingual Transformer Encoders 31 Dec 2020 · 0 repositories · arXiv:2012.15547
-
ECONET: Effective Continual Pretraining of Language Models for Event Temporal Reasoning 30 Dec 2020 · 2 repositories · arXiv:2012.15283
-
Deriving Contextualised Semantic Features from BERT (and Other Transformer Model) Embeddings 30 Dec 2020 · 0 repositories · arXiv:2012.15353
-
Improving BERT with Syntax-aware Local Attention 30 Dec 2020 · 1 repository · arXiv:2012.15150
-
MRI brain tumor segmentation and uncertainty estimation using 3D-UNet architectures 30 Dec 2020 · 1 repository · arXiv:2012.15294
-
Optimizing Deeper Transformers on Small Datasets 30 Dec 2020 · 1 repository · arXiv:2012.15355
-
Out of Order: How Important Is The Sequential Order of Words in a Sentence in Natural Language Understanding Tasks? 30 Dec 2020 · 0 repositories · arXiv:2012.15180
-
SemGloVe: Semantic Co-occurrences for GloVe from BERT 30 Dec 2020 · 3 repositories · arXiv:2012.15197
-
SkiNet: A Deep Learning Solution for Skin Lesion Diagnosis with Uncertainty Estimation and Explainability 30 Dec 2020 · 0 repositories · arXiv:2012.15049
-
Towards Unsupervised Deep Image Enhancement with Generative Adversarial Network 30 Dec 2020 · 1 repository · arXiv:2012.15020
-
Transformer for Image Quality Assessment 30 Dec 2020 · 0 repositories · arXiv:2101.01097
-
UnNatural Language Inference 30 Dec 2020 · 1 repository · arXiv:2101.00010
-
A Hierarchical Transformer with Speaker Modeling for Emotion Recognition in Conversation 29 Dec 2020 · 1 repository · arXiv:2012.14781
-
Kaleidoscope: An Efficient, Learnable Representation For All Structured Linear Maps 29 Dec 2020 · 2 repositories · arXiv:2012.14966Syntology official (archive's flag): 16 ran · 16 ran (of which 9 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 7 where Syntology's instrument failed) · 7 unverified (of 23 harvested samples)
-
LayoutLMv2: Multi-modal Pre-training for Visually-Rich Document Understanding 29 Dec 2020 · 9 repositories · arXiv:2012.14740
-
Robust Dialogue Utterance Rewriting as Sequence Tagging 29 Dec 2020 · 1 repository · arXiv:2012.14535
-
Code Summarization with Structure-induced Transformer 29 Dec 2020 · 1 repository · arXiv:2012.14710
-
A Paragraph-level Multi-task Learning Model for Scientific Fact-Verification 28 Dec 2020 · 1 repository · arXiv:2012.14500
-
BURT: BERT-inspired Universal Representation from Learning Meaningful Segment 28 Dec 2020 · 0 repositories · arXiv:2012.14320
-
Lattice-Free MMI Adaptation Of Self-Supervised Pretrained Acoustic Models 28 Dec 2020 · 2 repositories · arXiv:2012.14252
-
Red Dragon AI at TextGraphs 2020 Shared Task: LIT : LSTM-Interleaved Transformer for Multi-Hop Explanation Ranking 28 Dec 2020 · 1 repository · arXiv:2012.14164
-
Syntax-Enhanced Pre-trained Model 28 Dec 2020 · 1 repository · arXiv:2012.14116
-
Towards a category-extended object detector with limited data 28 Dec 2020 · 0 repositories · arXiv:2012.14115
-
TransPose: Keypoint Localization via Transformer 28 Dec 2020 · 1 repository · arXiv:2012.14214
-
A multi-task learning network using shared BERT models for aspect-based sentiment analysis 27 Dec 2020 · 0 repositories
-
ALP-KD: Attention-Based Layer Projection for Knowledge Distillation 27 Dec 2020 · 0 repositories · arXiv:2012.14022
-
An Embarrassingly Simple Model for Dialogue Relation Extraction 27 Dec 2020 · 1 repository · arXiv:2012.13873
-
Inserting Information Bottlenecks for Attribution in Transformers 27 Dec 2020 · 1 repository · arXiv:2012.13838
-
Learning Light-Weight Translation Models from Deep Transformer 27 Dec 2020 · 1 repository · arXiv:2012.13866Syntology official (archive's flag): 1 ran · 2 ran (of which 2 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples) · 1 pointer-only (licence)
-
MeDAL: Medical Abbreviation Disambiguation Dataset for Natural Language Understanding Pretraining 27 Dec 2020 · 1 repository · arXiv:2012.13978
-
Portfolio Optimization with 2D Relative-Attentional Gated Transformer 27 Dec 2020 · 0 repositories · arXiv:2101.03138
-
SG-Net: Syntax Guided Transformer for Language Representation 27 Dec 2020 · 0 repositories · arXiv:2012.13915
-
Bayesian Graph Neural Network for Fast identification of critical nodes in Uncertain Complex Networks 26 Dec 2020 · 0 repositories · arXiv:2012.15733