Methods › General › Regularization › Weight Decay › Papers, page 49
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 49 of 108: papers 4,801 to 4,900 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
System-Level Natural Language Feedback 23 Jun 2023 · 1 repository · arXiv:2306.13588Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale 23 Jun 2023 · 1 repository · arXiv:2306.15687Syntology 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 3 honoured, 3 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 10 harvested samples) · 3 pointer-only (licence)
-
Cross-lingual Cross-temporal Summarization: Dataset, Models, Evaluation 22 Jun 2023 · 1 repository · arXiv:2306.12916
-
Named entity recognition in resumes 22 Jun 2023 · 0 repositories · arXiv:2306.13062
-
Prompt to GPT-3: Step-by-Step Thinking Instructions for Humor Generation 22 Jun 2023 · 1 repository · arXiv:2306.13195
-
FlakyFix: Using Large Language Models for Predicting Flaky Test Fix Categories and Test Code Repair 21 Jun 2023 · 0 repositories · arXiv:2307.00012
-
Investigating Pre-trained Language Models on Cross-Domain Datasets, a Step Closer to General AI 21 Jun 2023 · 0 repositories · arXiv:2306.12205
-
Solving and Generating NPR Sunday Puzzles with Large Language Models 21 Jun 2023 · 1 repository · arXiv:2306.12255
-
Which Spurious Correlations Impact Reasoning in NLI Models? A Visual Interactive Diagnosis through Data-Constrained Counterfactuals 21 Jun 2023 · 0 repositories · arXiv:2306.12146
-
DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models 20 Jun 2023 · 0 repositories · arXiv:2306.11698
-
Event Stream GPT: A Data Pre-processing and Modeling Library for Generative, Pre-trained Transformers over Continuous-time Sequences of Complex Events 20 Jun 2023 · 1 repository · arXiv:2306.11547Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples)
-
InRank: Incremental Low-Rank Learning 20 Jun 2023 · 1 repository · arXiv:2306.11250
-
Learning to Generate Better Than Your LLM 20 Jun 2023 · 1 repository · arXiv:2306.11816
-
Safe, Efficient, Comfort, and Energy-saving Automated Driving through Roundabout Based on Deep Reinforcement Learning 20 Jun 2023 · 0 repositories · arXiv:2306.11465
-
Textbooks Are All You Need 20 Jun 2023 · 0 repositories · arXiv:2306.11644
-
A Preliminary Study of ChatGPT on News Recommendation: Personalization, Provider Fairness, Fake News 19 Jun 2023 · 1 repository · arXiv:2306.10702
-
BayLing: Bridging Cross-lingual Alignment and Instruction Following through Interactive Translation for Large Language Models 19 Jun 2023 · 1 repository · arXiv:2306.10968
-
Fine-Tuning Language Models for Scientific Writing Support 19 Jun 2023 · 1 repository · arXiv:2306.10974
-
Generative Sequential Recommendation with GPTRec 19 Jun 2023 · 0 repositories · arXiv:2306.11114
-
SynerGPT: In-Context Learning for Personalized Drug Synergy Prediction and Drug Design 19 Jun 2023 · 0 repositories · arXiv:2307.11694
-
Instant Soup: Cheap Pruning Ensembles in A Single Pass Can Draw Lottery Tickets from Large Models 18 Jun 2023 · 1 repository · arXiv:2306.10460Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Enhancing social network hate detection using back translation and GPT-3 augmentations during training and test-time 17 Jun 2023 · 1 repository
-
Is Self-Repair a Silver Bullet for Code Generation? 16 Jun 2023 · 1 repository · arXiv:2306.09896
-
GPT4 is Slightly Helpful for Peer-Review Assistance: A Pilot Study 16 Jun 2023 · 2 repositories · arXiv:2307.05492
-
Investigating Masking-based Data Generation in Language Models 16 Jun 2023 · 0 repositories · arXiv:2307.00008
-
Revealing the impact of social circumstances on the selection of cancer therapy through natural language processing of social work notes 16 Jun 2023 · 0 repositories · arXiv:2306.09877
-
BED: Bi-Encoder-Based Detectors for Out-of-Distribution Detection 15 Jun 2023 · 1 repository · arXiv:2306.08852
-
ChessGPT: Bridging Policy Learning and Language Modeling 15 Jun 2023 · 1 repository · arXiv:2306.09200Syntology official (archive's flag): 14 ran · 14 ran (of which 8 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 1 violated, 9 with no contract checked; 4 where Syntology's instrument failed) · 6 unverified (of 20 harvested samples)
-
Distillation Strategies for Discriminative Speech Recognition Rescoring 15 Jun 2023 · 0 repositories · arXiv:2306.09452
-
Explore, Establish, Exploit: Red Teaming Language Models from Scratch 15 Jun 2023 · 3 repositories · arXiv:2306.09442Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Exploring the MIT Mathematics and EECS Curriculum Using Large Language Models 15 Jun 2023 · 0 repositories · arXiv:2306.08997
-
Mapping Researcher Activity based on Publication Data by means of Transformers 15 Jun 2023 · 0 repositories · arXiv:2306.09049
-
SLAMB: Accelerated Large Batch Training with Sparse Communication 15 Jun 2023 · 1 repository
-
Stochastic Re-weighted Gradient Descent via Distributionally Robust Optimization 15 Jun 2023 · 0 repositories · arXiv:2306.09222
-
The pop song generator: designing an online course to teach collaborative, creative AI 15 Jun 2023 · 0 repositories · arXiv:2306.10069
-
Thrilled by Your Progress! Large Language Models (GPT-4) No Longer Struggle to Pass Assessments in Higher Education Programming Courses 15 Jun 2023 · 0 repositories · arXiv:2306.10073
-
A semantically enhanced dual encoder for aspect sentiment triplet extraction 14 Jun 2023 · 1 repository · arXiv:2306.08373
-
Assessing the Effectiveness of GPT-3 in Detecting False Political Statements: A Case Study on the LIAR Dataset 14 Jun 2023 · 1 repository · arXiv:2306.08190
-
Building a Corpus for Biomedical Relation Extraction of Species Mentions 14 Jun 2023 · 0 repositories · arXiv:2306.08403
-
Language models are not naysayers: An analysis of language models on negation benchmarks 14 Jun 2023 · 1 repository · arXiv:2306.08189
-
Noise Stability Optimization for Finding Flat Minima: A Hessian-based Regularization Approach 14 Jun 2023 · 1 repository · arXiv:2306.08553Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 15 harvested samples)
-
Towards AGI in Computer Vision: Lessons Learned from GPT and Large Language Models 14 Jun 2023 · 0 repositories · arXiv:2306.08641
-
World-to-Words: Grounded Open Vocabulary Acquisition through Fast Mapping in Vision-Language Models 14 Jun 2023 · 1 repository · arXiv:2306.08685Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
Enhancing Social Network Hate Detection Using Back Translation and GPT-3 Augmentations During Training and Test-Time 13 Jun 2023 · 1 repository
-
FLamE: Few-shot Learning from Natural Language Explanations 13 Jun 2023 · 0 repositories · arXiv:2306.08042
-
GEmo-CLAP: Gender-Attribute-Enhanced Contrastive Language-Audio Pretraining for Accurate Speech Emotion Recognition 13 Jun 2023 · 0 repositories · arXiv:2306.07848
-
Human-Like Intuitive Behavior and Reasoning Biases Emerged in Language Models -- and Disappeared in GPT-4 13 Jun 2023 · 0 repositories · arXiv:2306.07622
-
Improving Zero-Shot Detection of Low Prevalence Chest Pathologies using Domain Pre-trained Language Models 13 Jun 2023 · 1 repository · arXiv:2306.08000
-
Monolingual and Cross-Lingual Knowledge Transfer for Topic Classification 13 Jun 2023 · 0 repositories · arXiv:2306.07797
-
A Survey of Vision-Language Pre-training from the Lens of Multimodal Machine Translation 12 Jun 2023 · 0 repositories · arXiv:2306.07198
-
Imbalanced Multi-label Classification for Business-related Text with Moderately Large Label Spaces 12 Jun 2023 · 0 repositories · arXiv:2306.07046
-
Linear Classifier: An Often-Forgotten Baseline for Text Classification 12 Jun 2023 · 1 repository · arXiv:2306.07111Syntology official: no sample here; runs from other or unrecorded repositories · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples)
-
Multimodal Audio-textual Architecture for Robust Spoken Language Understanding 12 Jun 2023 · 0 repositories · arXiv:2306.06819
-
On the N-gram Approximation of Pre-trained Language Models 12 Jun 2023 · 0 repositories · arXiv:2306.06892
-
Recursion of Thought: A Divide-and-Conquer Approach to Multi-Context Reasoning with Language Models 12 Jun 2023 · 1 repository · arXiv:2306.06891
-
The BEA 2023 Shared Task on Generating AI Teacher Responses in Educational Dialogues 12 Jun 2023 · 0 repositories · arXiv:2306.06941
-
Waffling around for Performance: Visual Classification with Random Words and Broad Concepts 12 Jun 2023 · 2 repositories · arXiv:2306.07282Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
EaSyGuide : ESG Issue Identification Framework leveraging Abilities of Generative Large Language Models 11 Jun 2023 · 1 repository · arXiv:2306.06662
-
Inductive reasoning in humans and large language models 11 Jun 2023 · 1 repository · arXiv:2306.06548
-
On Minimizing the Impact of Dataset Shifts on Actionable Explanations 11 Jun 2023 · 0 repositories · arXiv:2306.06716
-
RoBERTweet: A BERT Language Model for Romanian Tweets 11 Jun 2023 · 0 repositories · arXiv:2306.06598
-
Enhancing Low Resource NER Using Assisting Language And Transfer Learning 10 Jun 2023 · 0 repositories · arXiv:2306.06477
-
Medical Data Augmentation via ChatGPT: A Case Study on Medication Identification and Medication Event Classification 10 Jun 2023 · 0 repositories · arXiv:2306.07297
-
COVER: A Heuristic Greedy Adversarial Attack on Prompt-based Learning in Language Models 9 Jun 2023 · 0 repositories · arXiv:2306.05659
-
End-to-End Neural Network Compression via ℓ₁/ℓ₂ Regularized Latency Surrogates 9 Jun 2023 · 0 repositories · arXiv:2306.05785
-
Exploring the Responses of Large Language Models to Beginner Programmers' Help Requests 9 Jun 2023 · 0 repositories · arXiv:2306.05715
-
GPT-Calls: Enhancing Call Segmentation and Tagging by Generating Synthetic Conversations via Large Language Models 9 Jun 2023 · 0 repositories · arXiv:2306.07941
-
Implementing BERT and fine-tuned RobertA to detect AI generated news by ChatGPT 9 Jun 2023 · 0 repositories · arXiv:2306.07401
-
Language Models Can Learn Exceptions to Syntactic Rules 9 Jun 2023 · 1 repository · arXiv:2306.05969
-
Prodigy: An Expeditiously Adaptive Parameter-Free Learner 9 Jun 2023 · 1 repository · arXiv:2306.06101
-
Reliability Check: An Analysis of GPT-3's Response to Sensitive Topics and Prompt Wording 9 Jun 2023 · 2 repositories · arXiv:2306.06199
-
Understanding Telecom Language Through Large Language Models 9 Jun 2023 · 0 repositories · arXiv:2306.07933
-
Augmenting Hessians with Inter-Layer Dependencies for Mixed-Precision Post-Training Quantization 8 Jun 2023 · 0 repositories · arXiv:2306.04879
-
Bias Against 93 Stigmatized Groups in Masked Language Models and Downstream Sentiment Classification Tasks 8 Jun 2023 · 1 repository · arXiv:2306.05550Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Extensive Evaluation of Transformer-based Architectures for Adverse Drug Events Extraction 8 Jun 2023 · 1 repository · arXiv:2306.05276
-
Leveraging Language Identification to Enhance Code-Mixed Text Classification 8 Jun 2023 · 0 repositories · arXiv:2306.04964
-
Mixture-of-Supernets: Improving Weight-Sharing Supernet Training with Architecture-Routed Mixture-of-Experts 8 Jun 2023 · 1 repository · arXiv:2306.04845
-
NOWJ at COLIEE 2023 -- Multi-Task and Ensemble Approaches in Legal Information Processing 8 Jun 2023 · 0 repositories · arXiv:2306.04903
-
PandaLM: An Automatic Evaluation Benchmark for LLM Instruction Tuning Optimization 8 Jun 2023 · 2 repositories · arXiv:2306.05087Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Prefer to Classify: Improving Text Classifiers via Auxiliary Preference Learning 8 Jun 2023 · 1 repository · arXiv:2306.04925Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Progression Cognition Reinforcement Learning with Prioritized Experience for Multi-Vehicle Pursuit 8 Jun 2023 · 1 repository · arXiv:2306.05016
-
The ADAIO System at the BEA-2023 Shared Task on Generating AI Teacher Responses in Educational Dialogues 8 Jun 2023 · 0 repositories · arXiv:2306.05360
-
ToolAlpaca: Generalized Tool Learning for Language Models with 3000 Simulated Cases 8 Jun 2023 · 3 repositories · arXiv:2306.05301Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples)
-
On the Detectability of ChatGPT Content: Benchmarking, Methodology, and Evaluation through the Lens of Academic Writing 7 Jun 2023 · 2 repositories · arXiv:2306.05524
-
Good Data, Large Data, or No Data? Comparing Three Approaches in Developing Research Aspect Classifiers for Biomedical Papers 7 Jun 2023 · 1 repository · arXiv:2306.04820
-
GPT Self-Supervision for a Better Data Annotator 7 Jun 2023 · 0 repositories · arXiv:2306.04349
-
Personality testing of Large Language Models: Limited temporal stability, but highlighted prosociality 7 Jun 2023 · 0 repositories · arXiv:2306.04308
-
ScienceBenchmark: A Complex Real-World Benchmark for Evaluating Natural Language to SQL Systems 7 Jun 2023 · 0 repositories · arXiv:2306.04743
-
The Two Word Test: A Semantic Benchmark for Large Language Models 7 Jun 2023 · 1 repository · arXiv:2306.04610
-
An Empirical Analysis of Parameter-Efficient Methods for Debiasing Pre-Trained Language Models 6 Jun 2023 · 1 repository · arXiv:2306.04067
-
Certified Deductive Reasoning with Language Models 6 Jun 2023 · 1 repository · arXiv:2306.04031Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Detecting Human Rights Violations on Social Media during Russia-Ukraine War 6 Jun 2023 · 0 repositories · arXiv:2306.05370
-
Iterative Translation Refinement with Large Language Models 6 Jun 2023 · 0 repositories · arXiv:2306.03856
-
Language acquisition: do children and language models follow similar learning stages? 6 Jun 2023 · 0 repositories · arXiv:2306.03586
-
LEACE: Perfect linear concept erasure in closed form 6 Jun 2023 · 2 repositories · arXiv:2306.03819Syntology official (archive's flag): 3 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples)
-
On the Difference of BERT-style and CLIP-style Text Encoders 6 Jun 2023 · 1 repository · arXiv:2306.03678
-
Analyzing Syntactic Generalization Capacity of Pre-trained Language Models on Japanese Honorific Conversion 5 Jun 2023 · 0 repositories · arXiv:2306.03055
-
ChatGPT as a mapping assistant: A novel method to enrich maps with generative AI and content derived from street-level photographs 5 Jun 2023 · 0 repositories · arXiv:2306.03204
-
COMET: Learning Cardinality Constrained Mixture of Experts with Trees and Local Search 5 Jun 2023 · 2 repositories · arXiv:2306.02824
-
Efficient GPT Model Pre-training using Tensor Train Matrix Representation 5 Jun 2023 · 0 repositories · arXiv:2306.02697