Methods › General › Regularization › Weight Decay › Papers, page 44
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 44 of 108: papers 4,301 to 4,400 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
An evaluation of GPT models for phenotype concept recognition 29 Sep 2023 · 0 repositories · arXiv:2309.17169
-
Benchmarking the Abilities of Large Language Models for RDF Knowledge Graph Creation and Comprehension: How Well Do LLMs Speak Turtle? 29 Sep 2023 · 3 repositories · arXiv:2309.17122
-
DyVal: Dynamic Evaluation of Large Language Models for Reasoning Tasks 29 Sep 2023 · 1 repository · arXiv:2309.17167
-
Intuitive or Dependent? Investigating LLMs' Behavior Style to Conflicting Prompts 29 Sep 2023 · 0 repositories · arXiv:2309.17415
-
Cooperation, Competition, and Maliciousness: LLM-Stakeholders Interactive Negotiation 29 Sep 2023 · 2 repositories · arXiv:2309.17234Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Revolutionizing Mobile Interaction: Enabling a 3 Billion Parameter GPT LLM on Mobile 29 Sep 2023 · 0 repositories · arXiv:2310.01434
-
Split and Merge: Aligning Position Biases in LLM-based Evaluators 29 Sep 2023 · 0 repositories · arXiv:2310.01432
-
Symmetry Induces Structure and Constraint of Learning 29 Sep 2023 · 0 repositories · arXiv:2309.16932
-
Training and inference of large language models using 8-bit floating point 29 Sep 2023 · 0 repositories · arXiv:2309.17224
-
AE-GPT: Using Large Language Models to Extract Adverse Events from Surveillance Reports-A Use Case with Influenza Vaccine Adverse Events 28 Sep 2023 · 0 repositories · arXiv:2309.16150
-
GPT-Fathom: Benchmarking Large Language Models to Decipher the Evolutionary Path towards GPT-4 and Beyond 28 Sep 2023 · 1 repository · arXiv:2309.16583Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples)
-
Hallucination Reduction in Long Input Text Summarization 28 Sep 2023 · 1 repository · arXiv:2309.16781
-
Large Language Model Soft Ideologization via AI-Self-Consciousness 28 Sep 2023 · 0 repositories · arXiv:2309.16167
-
Stress Testing Chain-of-Thought Prompting for Large Language Models 28 Sep 2023 · 0 repositories · arXiv:2309.16621
-
Deep Out-of-Distribution Uncertainty Quantification via Weight Entropy Maximization 27 Sep 2023 · 1 repository · arXiv:2309.15704
-
MKRAG: Medical Knowledge Retrieval Augmented Generation for Medical Question Answering 27 Sep 2023 · 0 repositories · arXiv:2309.16035
-
MindGPT: Interpreting What You See with Non-invasive Brain Recordings 27 Sep 2023 · 1 repository · arXiv:2309.15729Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
NLPBench: Evaluating Large Language Models on Solving NLP Problems 27 Sep 2023 · 1 repository · arXiv:2309.15630
-
An NLP Benchmark Dataset for Assessing Corporate Climate Policy Engagement 26 Sep 2023 · 0 repositories
-
CAPP-130: A Corpus of Chinese Application Privacy Policy Summarization and Interpretation 26 Sep 2023 · 1 repository
-
Legal Question-Answering in the Indian Context: Efficacy, Challenges, and Potential of Modern AI Models 26 Sep 2023 · 0 repositories · arXiv:2309.14735
-
Exploring Small Language Models with Prompt-Learning Paradigm for Efficient Domain-Specific Text Classification 26 Sep 2023 · 0 repositories · arXiv:2309.14779
-
How to Catch an AI Liar: Lie Detection in Black-Box LLMs by Asking Unrelated Questions 26 Sep 2023 · 1 repository · arXiv:2309.15840Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Low-rank Adaptation of Large Language Model Rescoring for Parameter-Efficient Speech Recognition 26 Sep 2023 · 0 repositories · arXiv:2309.15223
-
RAGAS: Automated Evaluation of Retrieval Augmented Generation 26 Sep 2023 · 3 repositories · arXiv:2309.15217Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
RankVicuna: Zero-Shot Listwise Document Reranking with Open-Source Large Language Models 26 Sep 2023 · 3 repositories · arXiv:2309.15088
-
Supersonic: Learning to Generate Source Code Optimizations in C/C++ 26 Sep 2023 · 1 repository · arXiv:2309.14846
-
Comprehensive Overview of Named Entity Recognition: Models, Domain-Specific Applications and Challenges 25 Sep 2023 · 0 repositories · arXiv:2309.14084
-
Enhancing data efficiency in reinforcement learning: a novel imagination mechanism based on mesh information propagation 25 Sep 2023 · 2 repositories · arXiv:2309.14243
-
Evaluating Cognitive Maps and Planning in Large Language Models with CogEval 25 Sep 2023 · 0 repositories · arXiv:2309.15129
-
LogGPT: Log Anomaly Detection via GPT 25 Sep 2023 · 1 repository · arXiv:2309.14482
-
Watch Your Language: Investigating Content Moderation with Large Language Models 25 Sep 2023 · 0 repositories · arXiv:2309.14517
-
Accelerating Large Batch Training via Gradient Signal to Noise Ratio (GSNR) 24 Sep 2023 · 0 repositories · arXiv:2309.13681
-
Does the "most sinfully decadent cake ever" taste good? Answering Yes/No Questions from Figurative Contexts 24 Sep 2023 · 0 repositories · arXiv:2309.13748
-
Seeing Is Not Always Believing: Invisible Collision Attack and Defence on Pre-Trained Models 24 Sep 2023 · 1 repository · arXiv:2309.13579
-
A Chat About Boring Problems: Studying GPT-based text normalization 23 Sep 2023 · 0 repositories · arXiv:2309.13426
-
Probing the Moral Development of Large Language Models through Defining Issues Test 23 Sep 2023 · 0 repositories · arXiv:2309.13356
-
Lexical Squad@Multimodal Hate Speech Event Detection 2023: Multimodal Hate Speech Detection using Fused Ensemble Approach 23 Sep 2023 · 1 repository · arXiv:2309.13354
-
AMPLIFY:Attention-based Mixup for Performance Improvement and Label Smoothing in Transformer 22 Sep 2023 · 1 repository · arXiv:2309.12689
-
BenLLMEval: A Comprehensive Evaluation into the Potentials and Pitfalls of Large Language Models on Bengali NLP 22 Sep 2023 · 0 repositories · arXiv:2309.13173
-
Contextual Emotion Estimation from Image Captions 22 Sep 2023 · 0 repositories · arXiv:2309.13136
-
Investigating Large Language Models and Control Mechanisms to Improve Text Readability of Biomedical Abstracts 22 Sep 2023 · 1 repository · arXiv:2309.13202
-
Large Language Models Are Also Good Prototypical Commonsense Reasoners 22 Sep 2023 · 0 repositories · arXiv:2309.13165
-
SPION: Layer-Wise Sparse Training of Transformer via Convolutional Flood Filling 22 Sep 2023 · 0 repositories · arXiv:2309.12578
-
TOPFORMER: Topology-Aware Authorship Attribution of Deepfake Texts with Diverse Writing Styles 22 Sep 2023 · 1 repository · arXiv:2309.12934
-
Goal-Oriented Prompt Attack and Safety Evaluation for LLMs 21 Sep 2023 · 2 repositories · arXiv:2309.11830
-
A Long N-step Surrogate Stage Reward for Deep Reinforcement Learning 21 Sep 2023 · 0 repositories
-
Bad Actor, Good Advisor: Exploring the Role of Large Language Models in Fake News Detection 21 Sep 2023 · 1 repository · arXiv:2309.12247
-
BayesTune: Bayesian Sparse Deep Model Fine-tuning 21 Sep 2023 · 1 repository
-
Constraints First: A New MDD-based Model to Generate Sentences Under Constraints 21 Sep 2023 · 0 repositories · arXiv:2309.12415
-
Deconstructing Data Reconstruction: Multiclass, Weight Decay and General Losses 21 Sep 2023 · 1 repository
-
Double Gumbel Q-Learning 21 Sep 2023 · 1 repository
-
FedNAR: Federated Optimization with Normalized Annealing Regularization 21 Sep 2023 · 1 repository
-
Implicit Differentiable Outlier Detection Enable Robust Deep Multimodal Analysis 21 Sep 2023 · 1 repository
-
Making Scalable Meta Learning Practical 21 Sep 2023 · 1 repository
-
Marich: A Query-efficient Distributionally Equivalent Model Extraction Attack 21 Sep 2023 · 1 repository
-
MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models 21 Sep 2023 · 1 repository · arXiv:2309.12284Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 14 where Syntology's instrument failed) · 7 unverified (of 22 harvested samples)
-
On the Relationship between Skill Neurons and Robustness in Prompt Tuning 21 Sep 2023 · 1 repository · arXiv:2309.12263
-
[Re] Exploring the Role of Grammar and Word Choice in Bias Toward African American English (AAE) in Hate Speech Classification 21 Sep 2023 · 0 repositories
-
Safe Hierarchical Reinforcement Learning for CubeSat Task Scheduling Based on Energy Consumption 21 Sep 2023 · 0 repositories · arXiv:2309.12004
-
SLHCat: Mapping Wikipedia Categories and Lists to DBpedia by Leveraging Semantic, Lexical, and Hierarchical Features 21 Sep 2023 · 0 repositories · arXiv:2309.11791
-
SPICED: News Similarity Detection Dataset with Multiple Topics and Complexity Levels 21 Sep 2023 · 0 repositories · arXiv:2309.13080
-
Stock Market Sentiment Classification and Backtesting via Fine-tuned BERT 21 Sep 2023 · 0 repositories · arXiv:2309.11979
-
TART: A plug-and-play Transformer module for task-agnostic reasoning 21 Sep 2023 · 1 repository
-
The Cambridge Law Corpus: A Dataset for Legal AI Research 21 Sep 2023 · 0 repositories · arXiv:2309.12269
-
The Reversal Curse: LLMs trained on "A is B" fail to learn "B is A" 21 Sep 2023 · 2 repositories · arXiv:2309.12288Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
TOA: Task-oriented Active VQA 21 Sep 2023 · 0 repositories
-
Towards Efficient Pre-Trained Language Model via Feature Correlation Distillation 21 Sep 2023 · 0 repositories
-
A Paradigm Shift in Machine Translation: Boosting Translation Performance of Large Language Models 20 Sep 2023 · 1 repository · arXiv:2309.11674Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Adaptive Multi-Agent Deep Reinforcement Learning for Timely Healthcare Interventions 20 Sep 2023 · 0 repositories · arXiv:2309.10980
-
AttentionMix: Data augmentation method that relies on BERT attention mechanism 20 Sep 2023 · 0 repositories · arXiv:2309.11104
-
Controlled Generation with Prompt Insertion for Natural Language Explanations in Grammatical Error Correction 20 Sep 2023 · 1 repository · arXiv:2309.11439
-
CoT-BERT: Enhancing Unsupervised Sentence Representation through Chain-of-Thought 20 Sep 2023 · 2 repositories · arXiv:2309.11143
-
Design of Chain-of-Thought in Math Problem Solving 20 Sep 2023 · 1 repository · arXiv:2309.11054
-
Fictional Worlds, Real Connections: Developing Community Storytelling Social Chatbots through LLMs 20 Sep 2023 · 0 repositories · arXiv:2309.11478
-
Generative AI in Mafia-like Game Simulation 20 Sep 2023 · 0 repositories · arXiv:2309.11672
-
GPT-MolBERTa: GPT Molecular Features Language Model for molecular property prediction 20 Sep 2023 · 0 repositories · arXiv:2310.03030
-
Safurai 001: New Qualitative Approach for Code LLM Evaluation 20 Sep 2023 · 1 repository · arXiv:2309.11385
-
Sequence-to-Sequence Spanish Pre-trained Language Models 20 Sep 2023 · 1 repository · arXiv:2309.11259
-
The Languini Kitchen: Enabling Language Modelling Research at Different Scales of Compute 20 Sep 2023 · 1 repository · arXiv:2309.11197
-
Language as the Medium: Multimodal Video Classification through text only 19 Sep 2023 · 0 repositories · arXiv:2309.10783
-
Mixed-Distil-BERT: Code-mixed Language Modeling for Bangla, English, and Hindi 19 Sep 2023 · 0 repositories · arXiv:2309.10272
-
Rigorously Assessing Natural Language Explanations of Neurons 19 Sep 2023 · 0 repositories · arXiv:2309.10312
-
Writer-Defined AI Personas for On-Demand Feedback Generation 19 Sep 2023 · 0 repositories · arXiv:2309.10433
-
Evaluation of GPT-3 for Anti-Cancer Drug Sensitivity Prediction 18 Sep 2023 · 0 repositories · arXiv:2309.10016
-
Facilitating NSFW Text Detection in Open-Domain Dialogue Systems via Knowledge Distillation 18 Sep 2023 · 1 repository · arXiv:2309.09749
-
Proposition from the Perspective of Chinese Language: A Chinese Proposition Classification Evaluation Benchmark 18 Sep 2023 · 0 repositories · arXiv:2309.09602
-
RECAP: Retrieval-Augmented Audio Captioning 18 Sep 2023 · 1 repository · arXiv:2309.09836Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Towards Ontology Construction with Language Models 18 Sep 2023 · 0 repositories · arXiv:2309.09898
-
Contrastive Decoding Improves Reasoning in Large Language Models 17 Sep 2023 · 0 repositories · arXiv:2309.09117
-
Detecting covariate drift in text data using document embeddings and dimensionality reduction 17 Sep 2023 · 1 repository · arXiv:2309.10000
-
Do Large GPT Models Discover Moral Dimensions in Language Representations? A Topological Study Of Sentence Embeddings 17 Sep 2023 · 0 repositories · arXiv:2309.09397
-
From Cooking Recipes to Robot Task Trees -- Improving Planning Correctness and Task Efficiency by Leveraging LLMs with a Knowledge Network 17 Sep 2023 · 0 repositories · arXiv:2309.09181
-
Decoder-only Architecture for Speech Recognition with CTC Prompts and Text Data Augmentation 16 Sep 2023 · 0 repositories · arXiv:2309.08876
-
Has Sentiment Returned to the Pre-pandemic Level? A Sentiment Analysis Using U.S. College Subreddit Data from 2019 to 2022 16 Sep 2023 · 1 repository · arXiv:2309.08845
-
Struc-Bench: Are Large Language Models Really Good at Generating Complex Structured Data? 16 Sep 2023 · 1 repository · arXiv:2309.08963
-
A Modern Turkish Poet: Fine-Tuned GPT-2 15 Sep 2023 · 1 repository
-
Advancing the Evaluation of Traditional Chinese Language Models: Towards a Comprehensive Benchmark Suite 15 Sep 2023 · 1 repository · arXiv:2309.08448
-
AlbNER: A Corpus for Named Entity Recognition in Albanian 15 Sep 2023 · 0 repositories · arXiv:2309.08741
-
Indian-BhED: A Dataset for Measuring India-Centric Biases in Large Language Models 15 Sep 2023 · 1 repository · arXiv:2309.08573