Methods › General › Regularization › Weight Decay › Papers, page 70
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 70 of 108: papers 6,901 to 7,000 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Learning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction 5 Jan 2022 · 2 repositories · arXiv:2201.02184
-
Comparison of biomedical relationship extraction methods and models for knowledge graph creation 5 Jan 2022 · 0 repositories · arXiv:2201.01647
-
Submix: Practical Private Prediction for Large-Scale Language Models 4 Jan 2022 · 0 repositories · arXiv:2201.00971
-
An Adversarial Benchmark for Fake News Detection Models 3 Jan 2022 · 1 repository · arXiv:2201.00912
-
3DPG: Distributed Deep Deterministic Policy Gradient Algorithms for Networked Multi-Agent Systems 3 Jan 2022 · 0 repositories · arXiv:2201.00570
-
Which Student is Best? A Comprehensive Knowledge Distillation Exam for Task-Specific BERT Models 3 Jan 2022 · 0 repositories · arXiv:2201.00558
-
On Sensitivity of Deep Learning Based Text Classification Algorithms to Practical Input Perturbations 2 Jan 2022 · 0 repositories · arXiv:2201.00318
-
Accelerating Neural Network Optimization Through an Automated Control Theory Lens 1 Jan 2022 · 0 repositories
-
Continual Stereo Matching of Continuous Driving Scenes With Growing Architecture 1 Jan 2022 · 1 repository
-
Expanding Large Pre-Trained Unimodal Models With Multimodal Information Injection for Image-Text Multimodal Classification 1 Jan 2022 · 0 repositories
-
SpaceEdit: Learning a Unified Editing Space for Open-Domain Image Color Editing 1 Jan 2022 · 0 repositories
-
Toward Pareto Efficient Fairness-Utility Trade-off inRecommendation through Reinforcement Learning 1 Jan 2022 · 0 repositories · arXiv:2201.00140
-
A Neural Network Solves, Explains, and Generates University Math Problems by Program Synthesis and Few-Shot Learning at Human Level 31 Dec 2021 · 1 repository · arXiv:2112.15594Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Clustering Vietnamese Conversations From Facebook Page To Build Training Dataset For Chatbot 31 Dec 2021 · 1 repository · arXiv:2112.15338
-
Automatic Mixed-Precision Quantization Search of BERT 30 Dec 2021 · 0 repositories · arXiv:2112.14938
-
EvoMoE: An Evolutional Mixture-of-Experts Training Framework via Dense-To-Sparse Gate 29 Dec 2021 · 2 repositories · arXiv:2112.14397Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
The University of Texas at Dallas HLTRI's Participation in EPIC-QA: Searching for Entailed Questions Revealing Novel Answer Nuggets 28 Dec 2021 · 0 repositories · arXiv:2112.13946
-
Contextual Sentence Analysis for the Sentiment Prediction on Financial Data 27 Dec 2021 · 0 repositories · arXiv:2112.13790
-
Event-based clinical findings extraction from radiology reports with pre-trained language model 27 Dec 2021 · 1 repository · arXiv:2112.13512
-
Mind the Gap: Cross-Lingual Information Retrieval with Hierarchical Knowledge Enhancement 27 Dec 2021 · 0 repositories · arXiv:2112.13510
-
Multi-Image Visual Question Answering 27 Dec 2021 · 1 repository · arXiv:2112.13706
-
Secondary Use of Clinical Problem List Entries for Neural Network-Based Disease Code Assignment 27 Dec 2021 · 0 repositories · arXiv:2112.13756
-
Evaluating Contextual Embeddings and their Extraction Layers for Depression Assessment 27 Dec 2021 · 0 repositories · arXiv:2112.13795
-
An Ensemble of Pre-trained Transformer Models For Imbalanced Multiclass Malware Classification 25 Dec 2021 · 1 repository · arXiv:2112.13236
-
CABACE: Injecting Character Sequence Information and Domain Knowledge for Enhanced Acronym and Long-Form Extraction 25 Dec 2021 · 1 repository · arXiv:2112.13237
-
Deeper Clinical Document Understanding Using Relation Extraction 25 Dec 2021 · 1 repository · arXiv:2112.13259
-
Distilling the Knowledge of Romanian BERTs Using Multiple Teachers 23 Dec 2021 · 1 repository · arXiv:2112.12650
-
ERNIE 3.0 Titan: Exploring Larger-scale Knowledge Enhanced Pre-training for Language Understanding and Generation 23 Dec 2021 · 3 repositories · arXiv:2112.12731
-
Adaptive Beam Search to Enhance On-device Abstractive Summarization 22 Dec 2021 · 0 repositories · arXiv:2201.02739
-
Consistency and Coherence from Points of Contextual Similarity 22 Dec 2021 · 0 repositories · arXiv:2112.11638
-
Deep Reinforcement Learning for Optimal Power Flow with Renewables Using Graph Information 22 Dec 2021 · 0 repositories · arXiv:2112.11461
-
DB-BERT: a Database Tuning Tool that "Reads the Manual" 21 Dec 2021 · 0 repositories · arXiv:2112.10925
-
Predicting Job Titles from Job Descriptions with Multi-label Text Classification 21 Dec 2021 · 1 repository · arXiv:2112.11052
-
Few-shot Learning with Multilingual Language Models 20 Dec 2021 · 2 repositories · arXiv:2112.10668
-
Training dataset and dictionary sizes matter in BERT models: the case of Baltic languages 20 Dec 2021 · 0 repositories · arXiv:2112.10553
-
Analysis and Mitigation of Dataset Artifacts in OpenAI GPT-3 19 Dec 2021 · 0 repositories
-
Data Augmentation for Mental Health Classification on Social Media 19 Dec 2021 · 0 repositories · arXiv:2112.10064
-
Leveraging Transformers for Hate Speech Detection in Conversational Code-Mixed Tweets 18 Dec 2021 · 0 repositories · arXiv:2112.09986
-
Zero-shot and Few-shot Learning with Knowledge Graphs: A Comprehensive Survey 18 Dec 2021 · 0 repositories · arXiv:2112.10006
-
Syntactic-GCN Bert based Chinese Event Extraction 18 Dec 2021 · 0 repositories · arXiv:2112.09939
-
A High-Precision Health-relatedness Score for Phrases to Mine Cause-Effect Statements from the Web 17 Dec 2021 · 0 repositories
-
Can Machine Learning Tools Support the Identification of Sustainable Design Leads From Product Reviews? Opportunities and Challenges 17 Dec 2021 · 0 repositories · arXiv:2112.09391
-
Challenging America: Modeling language in longer time scales 17 Dec 2021 · 0 repositories
-
Explain, Edit, and Understand: Rethinking User Study Design for Evaluating Model Explanations 17 Dec 2021 · 1 repository · arXiv:2112.09669
-
Joint Chinese Word Segmentation and Part-of-speech Tagging via Two-stage Span Labeling 17 Dec 2021 · 0 repositories · arXiv:2112.09488
-
Learning to Win Lottery Tickets in BERT Transfer via Task-agnostic Mask Training 17 Dec 2021 · 0 repositories
-
Rank4Class: A Ranking Formulation for Multiclass Classification 17 Dec 2021 · 0 repositories · arXiv:2112.09727
-
Towards Faithful Personalized Response Selection in Retrieval Based Dialog Systems 17 Dec 2021 · 0 repositories
-
WebGPT: Browser-assisted question-answering with human feedback 17 Dec 2021 · 2 repositories · arXiv:2112.09332
-
An Empirical Study on Transfer Learning for Privilege Review 16 Dec 2021 · 0 repositories · arXiv:2112.08606
-
Call for Customized Conversation: Customized Conversation Grounding Persona and Knowledge 16 Dec 2021 · 3 repositories · arXiv:2112.08619
-
Knowledge-Augmented Language Models for Cause-Effect Relation Classification 16 Dec 2021 · 1 repository · arXiv:2112.08615Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Does Pre-training Induce Systematic Inference? How Masked Language Models Acquire Commonsense Knowledge 16 Dec 2021 · 0 repositories · arXiv:2112.08583
-
Efficient Hierarchical Domain Adaptation for Pretrained Language Models 16 Dec 2021 · 1 repository · arXiv:2112.08786
-
Few-Shot Semantic Parsing with Language Models Trained On Code 16 Dec 2021 · 0 repositories · arXiv:2112.08696
-
Reconsidering the Past: Optimizing Hidden States in Language Models 16 Dec 2021 · 0 repositories · arXiv:2112.08653
-
Reframing Human-AI Collaboration for Generating Free-Text Explanations 16 Dec 2021 · 1 repository · arXiv:2112.08674Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
3D Question Answering 15 Dec 2021 · 0 repositories · arXiv:2112.08359
-
Applying SoftTriple Loss for Supervised Language Model Fine Tuning 15 Dec 2021 · 0 repositories · arXiv:2112.08462
-
Fine-Tuning Large Neural Language Models for Biomedical Natural Language Processing 15 Dec 2021 · 0 repositories · arXiv:2112.07869
-
One size does not fit all: Investigating strategies for differentially-private learning across NLP tasks 15 Dec 2021 · 1 repository · arXiv:2112.08159Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
One System to Rule them All: a Universal Intent Recognition System for Customer Service Chatbots 15 Dec 2021 · 0 repositories · arXiv:2112.08261
-
Tracing Text Provenance via Context-Aware Lexical Substitution 15 Dec 2021 · 0 repositories · arXiv:2112.07873
-
ACE-BERT: Adversarial Cross-modal Enhanced BERT for E-commerce Retrieval 14 Dec 2021 · 0 repositories · arXiv:2112.07209
-
Building on Huang et al. GlossBERT for Word Sense Disambiguation 14 Dec 2021 · 0 repositories · arXiv:2112.07089
-
Classifying Emails into Human vs Machine Category 14 Dec 2021 · 0 repositories · arXiv:2112.07742
-
CoCo-BERT: Improving Video-Language Pre-training with Contrastive Cross-modal Matching and Denoising 14 Dec 2021 · 0 repositories · arXiv:2112.07515
-
Epigenomic language models powered by Cerebras 14 Dec 2021 · 0 repositories · arXiv:2112.07571
-
From Dense to Sparse: Contrastive Pruning for Better Pre-trained Language Model Compression 14 Dec 2021 · 2 repositories · arXiv:2112.07198
-
Measuring Fairness with Biased Rulers: A Survey on Quantifying Biases in Pretrained Language Models 14 Dec 2021 · 1 repository · arXiv:2112.07447
-
Text Classification Models for Form Entity Linking 14 Dec 2021 · 1 repository · arXiv:2112.07443
-
Towards a Unified Foundation Model: Jointly Pre-Training Transformers on Unpaired Images and Text 14 Dec 2021 · 0 repositories · arXiv:2112.07074
-
A Study on Token Pruning for ColBERT 13 Dec 2021 · 0 repositories · arXiv:2112.06540
-
Measuring Context-Word Biases in Lexical Semantic Datasets 13 Dec 2021 · 0 repositories · arXiv:2112.06733
-
Do Data-based Curricula Work? 13 Dec 2021 · 0 repositories · arXiv:2112.06510
-
Embracing Single Stride 3D Object Detector with Sparse Transformer 13 Dec 2021 · 2 repositories · arXiv:2112.06375
-
GLaM: Efficient Scaling of Language Models with Mixture-of-Experts 13 Dec 2021 · 0 repositories · arXiv:2112.06905
-
Keyphrase Generation Beyond the Boundaries of Title and Abstract 13 Dec 2021 · 1 repository · arXiv:2112.06776
-
Roof-Transformer: Divided and Joined Understanding with Knowledge Enhancement 13 Dec 2021 · 0 repositories · arXiv:2112.06736
-
WECHSEL: Effective initialization of subword embeddings for cross-lingual transfer of monolingual language models 13 Dec 2021 · 1 repository · arXiv:2112.06598Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Improving Logical-Level Natural Language Generation with Topic-Conditioned Data Augmentation and Logical Form Generation 12 Dec 2021 · 0 repositories · arXiv:2112.06240
-
Findings on Conversation Disentanglement 10 Dec 2021 · 0 repositories · arXiv:2112.05346
-
Multimodal Interactions Using Pretrained Unimodal Models for SIMMC 2.0 10 Dec 2021 · 1 repository · arXiv:2112.05328
-
Detecting potentially harmful and protective suicide-related content on twitter: A machine learning approach 9 Dec 2021 · 2 repositories · arXiv:2112.04796
-
From Scattered Sources to Comprehensive Technology Landscape: A Recommendation-based Retrieval Approach 9 Dec 2021 · 0 repositories · arXiv:2112.04810
-
Semantic Search as Extractive Paraphrase Span Detection 9 Dec 2021 · 1 repository · arXiv:2112.04886
-
Improving language models by retrieving from trillions of tokens 8 Dec 2021 · 2 repositories · arXiv:2112.04426Syntology 16 ran (of which 5 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 3 violated, 11 with no contract checked; 2 where Syntology's instrument failed) · 7 unverified (of 23 harvested samples) · 3 pointer-only (licence)
-
JABER and SABER: Junior and Senior Arabic BERt 8 Dec 2021 · 1 repository · arXiv:2112.04329
-
A Transferable Approach for Partitioning Machine Learning Models on Multi-Chip-Modules 7 Dec 2021 · 0 repositories · arXiv:2112.04041
-
raceBERT -- A Transformer-based Model for Predicting Race and Ethnicity from Names 7 Dec 2021 · 1 repository · arXiv:2112.03807
-
BERTMap: A BERT-based Ontology Alignment System 5 Dec 2021 · 1 repository · arXiv:2112.02682
-
Causal Distillation for Language Models 5 Dec 2021 · 1 repository · arXiv:2112.02505Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
DIBERT: Dependency Injected Bidirectional Encoder Representations from Transformers 5 Dec 2021 · 1 repository
-
Gaudí: Conversational Interactions with Deep Representations to Generate Image Collections 5 Dec 2021 · 0 repositories · arXiv:2112.04404
-
VarCLR: Variable Semantic Representation Pre-training via Contrastive Learning 5 Dec 2021 · 1 repository · arXiv:2112.02650Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Bridging Pre-trained Models and Downstream Tasks for Source Code Understanding 4 Dec 2021 · 1 repository · arXiv:2112.02268Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 13 harvested samples)
-
Representation Learning for Conversational Data using Discourse Mutual Information Maximization 4 Dec 2021 · 0 repositories · arXiv:2112.05787
-
Unraveling Social Perceptions & Behaviors towards Migrants on Twitter 4 Dec 2021 · 0 repositories · arXiv:2112.06642
-
A Novel Deep Parallel Time-series Relation Network for Fault Diagnosis 3 Dec 2021 · 0 repositories · arXiv:2112.03405
-
Augmenting Customer Support with an NLP-based Receptionist 3 Dec 2021 · 0 repositories · arXiv:2112.01959