Methods › General › Regularization › Weight Decay › Papers, page 36
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 36 of 108: papers 3,501 to 3,600 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
LLM Agents in Interaction: Measuring Personality Consistency and Linguistic Alignment in Interacting Populations of Large Language Models 5 Feb 2024 · 1 repository · arXiv:2402.02896
-
Multi-Lingual Malaysian Embedding: Leveraging Large Language Models for Semantic Representations 5 Feb 2024 · 0 repositories · arXiv:2402.03053
-
SWAG: Storytelling With Action Guidance 5 Feb 2024 · 1 repository · arXiv:2402.03483
-
UniMem: Towards a Unified View of Long-Context Large Language Models 5 Feb 2024 · 1 repository · arXiv:2402.03009
-
A Graph is Worth K Words: Euclideanizing Graph using Pure Transformer 4 Feb 2024 · 1 repository · arXiv:2402.02464Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
AutoTimes: Autoregressive Time Series Forecasters via Large Language Models 4 Feb 2024 · 1 repository · arXiv:2402.02370
-
Breaking MLPerf Training: A Case Study on Optimizing BERT 4 Feb 2024 · 0 repositories · arXiv:2402.02447
-
GeReA: Question-Aware Prompt Captions for Knowledge-based Visual Question Answering 4 Feb 2024 · 1 repository · arXiv:2402.02503Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
Improving Assessment of Tutoring Practices using Retrieval-Augmented Generation 4 Feb 2024 · 0 repositories · arXiv:2402.14594
-
Data Quality Matters: Suicide Intention Detection on Social Media Posts Using RoBERTa-CNN 3 Feb 2024 · 0 repositories · arXiv:2402.02262
-
DE³-BERT: Distance-Enhanced Early Exiting for BERT based on Prototypical Networks 3 Feb 2024 · 0 repositories · arXiv:2402.05948
-
EffiBench: Benchmarking the Efficiency of Automatically Generated Code 3 Feb 2024 · 1 repository · arXiv:2402.02037Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
MinMaxMin Q-learning 3 Feb 2024 · 0 repositories · arXiv:2402.05951
-
SQT -- std Q-target 3 Feb 2024 · 0 repositories · arXiv:2402.05950
-
Hierarchical Structure Enhances the Convergence and Generalizability of Linear Molecular Representation 3 Feb 2024 · 1 repository · arXiv:2402.02164
-
Clarifying the Path to User Satisfaction: An Investigation into Clarification Usefulness 2 Feb 2024 · 1 repository · arXiv:2402.01934
-
COMET: Generating Commit Messages using Delta Graph Context Representation 2 Feb 2024 · 0 repositories · arXiv:2402.01841
-
Can LLMs perform structured graph reasoning? 2 Feb 2024 · 1 repository · arXiv:2402.01805
-
Improving Sequential Recommendations with LLMs 2 Feb 2024 · 1 repository · arXiv:2402.01339Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
LLM-Detector: Improving AI-Generated Chinese Text Detection with Open-Source LLM Instruction Tuning 2 Feb 2024 · 1 repository · arXiv:2402.01158
-
Predicting ATP binding sites in protein sequences using Deep Learning and Natural Language Processing 2 Feb 2024 · 0 repositories · arXiv:2402.01829
-
Retrieval Augmented End-to-End Spoken Dialog Models 2 Feb 2024 · 0 repositories · arXiv:2402.01828
-
CorpusLM: Towards a Unified Language Model on Corpus for Knowledge-Intensive Tasks 2 Feb 2024 · 0 repositories · arXiv:2402.01176
-
Generation, Distillation and Evaluation of Motivational Interviewing-Style Reflections with a Foundational Language Model 1 Feb 2024 · 0 repositories · arXiv:2402.01051
-
HiQA: A Hierarchical Contextual Augmentation RAG for Multi-Documents QA 1 Feb 2024 · 0 repositories · arXiv:2402.01767
-
Learning Planning-based Reasoning by Trajectories Collection and Process Reward Synthesizing 1 Feb 2024 · 0 repositories · arXiv:2402.00658
-
ReAGent: A Model-agnostic Feature Attribution Method for Generative Language Models 1 Feb 2024 · 1 repository · arXiv:2402.00794
-
Self-Supervised Contrastive Pre-Training for Multivariate Point Processes 1 Feb 2024 · 0 repositories · arXiv:2402.00987
-
SPARQL Generation with Entity Pre-trained GPT for KG Question Answering 1 Feb 2024 · 1 repository · arXiv:2402.00969
-
Tiny Titans: Can Smaller Large Language Models Punch Above Their Weight in the Real World for Meeting Summarization? 1 Feb 2024 · 0 repositories · arXiv:2402.00841
-
Human-mediated Large Language Models for Robotic Intervention in Children with Autism Spectrum Disorders 1 Feb 2024 · 0 repositories · arXiv:2402.00260
-
ConSmax: Hardware-Friendly Alternative Softmax with Learnable Parameters 31 Jan 2024 · 1 repository · arXiv:2402.10930
-
Global-Liar: Factuality of LLMs over Time and Geographic Regions 31 Jan 2024 · 0 repositories · arXiv:2401.17839
-
Making a Long Story Short in Conversation Modeling 31 Jan 2024 · 0 repositories · arXiv:2402.00143
-
Mitigating the Influence of Distractor Tasks in LMs with Prior-Aware Decoding 31 Jan 2024 · 0 repositories · arXiv:2401.17692
-
Paramanu: A Family of Novel Efficient Generative Foundation Language Models for Indian Languages 31 Jan 2024 · 0 repositories · arXiv:2401.18034
-
RAG-Fusion: a New Take on Retrieval-Augmented Generation 31 Jan 2024 · 0 repositories · arXiv:2402.03367
-
Real Sparks of Artificial Intelligence and the Importance of Inner Interpretability 31 Jan 2024 · 0 repositories · arXiv:2402.00901
-
Uncertainty-Aware Explainable Recommendation with Large Language Models 31 Jan 2024 · 0 repositories · arXiv:2402.03366
-
A Preliminary Study on Using Large Language Models in Software Pentesting 30 Jan 2024 · 0 repositories · arXiv:2401.17459
-
Arabic Tweet Act: A Weighted Ensemble Pre-Trained Transformer Model for Classifying Arabic Speech Acts on Twitter 30 Jan 2024 · 0 repositories · arXiv:2401.17373
-
Breaking Free Transformer Models: Task-specific Context Attribution Promises Improved Generalizability Without Fine-tuning Pre-trained LLMs 30 Jan 2024 · 1 repository · arXiv:2401.16638
-
CRUD-RAG: A Comprehensive Chinese Benchmark for Retrieval-Augmented Generation of Large Language Models 30 Jan 2024 · 1 repository · arXiv:2401.17043Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Detecting mental disorder on social media: a ChatGPT-augmented explainable approach 30 Jan 2024 · 1 repository · arXiv:2401.17477
-
Detecting Racist Text in Bengali: An Ensemble Deep Learning Framework 30 Jan 2024 · 0 repositories · arXiv:2401.16748
-
Fine-tuning Transformer-based Encoder for Turkish Language Understanding Tasks 30 Jan 2024 · 0 repositories · arXiv:2401.17396
-
Large Multi-Modal Models (LMMs) as Universal Foundation Models for AI-Native Wireless Systems 30 Jan 2024 · 0 repositories · arXiv:2402.01748
-
LLaMP: Large Language Model Made Powerful for High-fidelity Materials Knowledge Retrieval and Distillation 30 Jan 2024 · 1 repository · arXiv:2401.17244Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models 30 Jan 2024 · 1 repository · arXiv:2401.16745Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 2 pointer-only (licence)
-
Single Word Change is All You Need: Designing Attacks and Defenses for Text Classifiers 30 Jan 2024 · 0 repositories · arXiv:2401.17196
-
Towards Generating Informative Textual Description for Neurons in Language Models 30 Jan 2024 · 0 repositories · arXiv:2401.16731
-
Credit Risk Meets Large Language Models: Building a Risk Indicator from Loan Descriptions in P2P Lending 29 Jan 2024 · 0 repositories · arXiv:2401.16458
-
Deep Reinforcement Learning for Voltage Control and Renewable Accommodation Using Spatial-Temporal Graph Information 29 Jan 2024 · 0 repositories · arXiv:2401.15848
-
Development and Testing of a Novel Large Language Model-Based Clinical Decision Support Systems for Medication Safety in 12 Clinical Specialties 29 Jan 2024 · 0 repositories · arXiv:2402.01741
-
Development and Testing of Retrieval Augmented Generation in Large Language Models -- A Case Study Report 29 Jan 2024 · 0 repositories · arXiv:2402.01733
-
Diverse, but Divisive: LLMs Can Exaggerate Gender Differences in Opinion Related to Harms of Misinformation 29 Jan 2024 · 0 repositories · arXiv:2401.16558
-
BPDec: Unveiling the Potential of Masked Language Modeling Decoder in BERT pretraining 29 Jan 2024 · 0 repositories · arXiv:2401.15861
-
E-EVAL: A Comprehensive Chinese K-12 Education Evaluation Benchmark for Large Language Models 29 Jan 2024 · 1 repository · arXiv:2401.15927Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Leveraging Professional Radiologists' Expertise to Enhance LLMs' Evaluation for Radiology Reports 29 Jan 2024 · 0 repositories · arXiv:2401.16578
-
LLM4Vuln: A Unified Evaluation Framework for Decoupling and Enhancing LLMs' Vulnerability Reasoning 29 Jan 2024 · 0 repositories · arXiv:2401.16185
-
Multi-class Regret Detection in Hindi Devanagari Script 29 Jan 2024 · 0 repositories · arXiv:2401.16561
-
ReGAL: Refactoring Programs to Discover Generalizable Abstractions 29 Jan 2024 · 1 repository · arXiv:2401.16467Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
An Insight into Security Code Review with LLMs: Capabilities, Obstacles, and Influential Factors 29 Jan 2024 · 0 repositories · arXiv:2401.16310
-
TrackGPT -- A generative pre-trained transformer for cross-domain entity trajectory forecasting 29 Jan 2024 · 0 repositories · arXiv:2402.00066
-
Contrastive Learning and Mixture of Experts Enables Precise Vector Embeddings 28 Jan 2024 · 1 repository · arXiv:2401.15713
-
UnMASKed: Quantifying Gender Biases in Masked Language Models through Linguistically Informed Job Market Prompts 28 Jan 2024 · 0 repositories · arXiv:2401.15798
-
ConvoSense: Overcoming Monotonous Commonsense Inferences for Conversational AI 27 Jan 2024 · 1 repository · arXiv:2401.15471
-
Enhancing Large Language Model Performance To Answer Questions and Extract Information More Accurately 27 Jan 2024 · 0 repositories · arXiv:2402.01722
-
Equipping Language Models with Tool Use Capability for Tabular Data Analysis in Finance 27 Jan 2024 · 0 repositories · arXiv:2401.15328
-
Fortifying Ethical Boundaries in AI: Advanced Strategies for Enhancing Security in Large Language Models 27 Jan 2024 · 0 repositories · arXiv:2402.01725
-
MultiHop-RAG: Benchmarking Retrieval-Augmented Generation for Multi-Hop Queries 27 Jan 2024 · 2 repositories · arXiv:2401.15391Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
From RAG to QA-RAG: Integrating Generative AI for Pharmaceutical Regulatory Compliance Process 26 Jan 2024 · 1 repository · arXiv:2402.01717
-
GeoDecoder: Empowering Multimodal Map Understanding 26 Jan 2024 · 0 repositories · arXiv:2401.15118
-
Scalable Qualitative Coding with LLMs: Chain-of-Thought Reasoning Matches Human Performance in Some Hermeneutic Tasks 26 Jan 2024 · 0 repositories · arXiv:2401.15170
-
The Power of Noise: Redefining Retrieval for RAG Systems 26 Jan 2024 · 3 repositories · arXiv:2401.14887Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
A comparative study of zero-shot inference with large language models and supervised modeling in breast cancer pathology classification 25 Jan 2024 · 0 repositories · arXiv:2401.13887
-
(Chat)GPT v BERT: Dawn of Justice for Semantic Change Detection 25 Jan 2024 · 1 repository · arXiv:2401.14040
-
DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence 25 Jan 2024 · 1 repository · arXiv:2401.14196Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Enhanced Labeling Technique for Reddit Text and Fine-Tuned Longformer Models for Classifying Depression Severity in English and Luganda 25 Jan 2024 · 0 repositories · arXiv:2401.14240
-
Evaluating GPT-3.5's Awareness and Summarization Abilities for European Constitutional Texts with Shared Topics 25 Jan 2024 · 0 repositories · arXiv:2401.14524
-
Investigate-Consolidate-Exploit: A General Strategy for Inter-Task Agent Self-Evolution 25 Jan 2024 · 0 repositories · arXiv:2401.13996
-
LongHealth: A Question Answering Benchmark with Long Clinical Documents 25 Jan 2024 · 1 repository · arXiv:2401.14490Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Socially Aware Synthetic Data Generation for Suicidal Ideation Detection Using Large Language Models 25 Jan 2024 · 0 repositories · arXiv:2402.01712
-
TrICy: Trigger-guided Data-to-text Generation with Intent aware Attention-Copy 25 Jan 2024 · 0 repositories · arXiv:2402.01714
-
Unmasking and Quantifying Racial Bias of Large Language Models in Medical Report Generation 25 Jan 2024 · 0 repositories · arXiv:2401.13867
-
ZS4C: Zero-Shot Synthesis of Compilable Code for Incomplete Code Snippets using LLMs 25 Jan 2024 · 0 repositories · arXiv:2401.14279
-
A Unified Approach to Emotion Detection and Task-Oriented Dialogue Modeling 24 Jan 2024 · 1 repository · arXiv:2401.13789
-
Automated Root Causing of Cloud Incidents using In-Context Learning with GPT-4 24 Jan 2024 · 0 repositories · arXiv:2401.13810
-
Can GPT-3.5 Generate and Code Discharge Summaries? 24 Jan 2024 · 1 repository · arXiv:2401.13512Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Discovering Mathematical Formulas from Data via GPT-guided Monte Carlo Tree Search 24 Jan 2024 · 0 repositories · arXiv:2401.14424
-
Evaluation of General Large Language Models in Contextually Assessing Semantic Concepts Extracted from Adult Critical Care Electronic Health Record Notes 24 Jan 2024 · 0 repositories · arXiv:2401.13588
-
Graph Guided Question Answer Generation for Procedural Question-Answering 24 Jan 2024 · 0 repositories · arXiv:2401.13594
-
How Good is ChatGPT at Face Biometrics? A First Look into Recognition, Soft Biometrics, and Explainability 24 Jan 2024 · 1 repository · arXiv:2401.13641
-
Proactive Emotion Tracker: AI-Driven Continuous Mood and Emotion Monitoring 24 Jan 2024 · 0 repositories · arXiv:2401.13722
-
Segment Any Cell: A SAM-based Auto-prompting Fine-tuning Framework for Nuclei Segmentation 24 Jan 2024 · 0 repositories · arXiv:2401.13220
-
Contrastive Learning in Distilled Models 23 Jan 2024 · 1 repository · arXiv:2401.12472
-
Fast Adversarial Training against Textual Adversarial Attacks 23 Jan 2024 · 0 repositories · arXiv:2401.12461
-
KAM-CoT: Knowledge Augmented Multimodal Chain-of-Thoughts Reasoning 23 Jan 2024 · 0 repositories · arXiv:2401.12863
-
Revolutionizing Retrieval-Augmented Generation with Enhanced PDF Structure Recognition 23 Jan 2024 · 0 repositories · arXiv:2401.12599
-
TroVE: Inducing Verifiable and Efficient Toolboxes for Solving Programmatic Tasks 23 Jan 2024 · 1 repository · arXiv:2401.12869Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 19 harvested samples) · 19 pointer-only (licence)