Methods › General › Regularization › Weight Decay › Papers, page 23
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 23 of 108: papers 2,201 to 2,300 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
TAGIFY: LLM-powered Tagging Interface for Improved Data Findability on OGD portals 26 Jul 2024 · 0 repositories · arXiv:2407.18764
-
The Cross-environment Hyperparameter Setting Benchmark for Reinforcement Learning 26 Jul 2024 · 0 repositories · arXiv:2407.18840
-
Using Large Language Models for the Interpretation of Building Regulations 26 Jul 2024 · 0 repositories · arXiv:2407.21060
-
Banyan: Improved Representation Learning with Explicit Structure 25 Jul 2024 · 0 repositories · arXiv:2407.17771
-
Closing the gap between open-source and commercial large language models for medical evidence summarization 25 Jul 2024 · 0 repositories · arXiv:2408.00588
-
Cost-effective Instruction Learning for Pathology Vision and Language Analysis 25 Jul 2024 · 1 repository · arXiv:2407.17734
-
PEFT-U: Parameter-Efficient Fine-Tuning for User Personalization 25 Jul 2024 · 1 repository · arXiv:2407.18078Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
PersonaGym: Evaluating Persona Agents and LLMs 25 Jul 2024 · 1 repository · arXiv:2407.18416
-
RoBERTa, ResNeXt and BiLSTM with self-attention: The ultimate trio for customer sentiment analysis 25 Jul 2024 · 0 repositories
-
The Geometry of Queries: Query-Based Innovations in Retrieval-Augmented Generation 25 Jul 2024 · 0 repositories · arXiv:2407.18044
-
Understanding the Interplay of Scale, Data, and Bias in Language Models: A Case Study with BERT 25 Jul 2024 · 0 repositories · arXiv:2407.21058
-
What Matters in Explanations: Towards Explainable Fake Review Detection Focusing on Transformers 24 Jul 2024 · 0 repositories · arXiv:2407.21056
-
A Comprehensive Approach to Misspelling Correction with BERT and Levenshtein Distance 24 Jul 2024 · 0 repositories · arXiv:2407.17383
-
A Novel Two-Step Fine-Tuning Pipeline for Cold-Start Active Learning in Text Classification Tasks 24 Jul 2024 · 0 repositories · arXiv:2407.17284
-
Bailicai: A Domain-Optimized Retrieval-Augmented Generation Framework for Medical Applications 24 Jul 2024 · 0 repositories · arXiv:2407.21055
-
Testing Large Language Models on Driving Theory Knowledge and Skills for Connected Autonomous Vehicles 24 Jul 2024 · 0 repositories · arXiv:2407.17211
-
Analyzing Polysemy Evolution Using Semantic Cells 23 Jul 2024 · 0 repositories · arXiv:2407.16110
-
Artificial Intelligence in Extracting Diagnostic Data from Dental Records 23 Jul 2024 · 0 repositories · arXiv:2407.21050
-
Data Mixture Inference: What do BPE Tokenizers Reveal about their Training Data? 23 Jul 2024 · 1 repository · arXiv:2407.16607Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Enhancing LLM's Cognition via Structurization 23 Jul 2024 · 1 repository · arXiv:2407.16434Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
LawLuo: A Multi-Agent Collaborative Framework for Multi-Round Chinese Legal Consultation 23 Jul 2024 · 0 repositories · arXiv:2407.16252
-
Patched RTC: evaluating LLMs for diverse software development tasks 23 Jul 2024 · 1 repository · arXiv:2407.16557
-
Retrieval Augmented Generation or Long-Context LLMs? A Comprehensive Study and Hybrid Approach 23 Jul 2024 · 0 repositories · arXiv:2407.16833
-
Robust Privacy Amidst Innovation with Large Language Models Through a Critical Assessment of the Risks 23 Jul 2024 · 1 repository · arXiv:2407.16166
-
TookaBERT: A Step Forward for Persian NLU 23 Jul 2024 · 0 repositories · arXiv:2407.16382
-
An Empirical Comparison of Video Frame Sampling Methods for Multi-Modal RAG Retrieval 22 Jul 2024 · 0 repositories · arXiv:2408.03340
-
Customized Retrieval Augmented Generation and Benchmarking for EDA Tool Documentation QA 22 Jul 2024 · 0 repositories · arXiv:2407.15353
-
Impacts of Anthropomorphizing Large Language Models in Learning Environments 22 Jul 2024 · 0 repositories · arXiv:2408.03945
-
Imposter.AI: Adversarial Attacks with Hidden Intentions towards Aligned Large Language Models 22 Jul 2024 · 0 repositories · arXiv:2407.15399
-
Inverted Activations: Reducing Memory Footprint in Neural Network Training 22 Jul 2024 · 1 repository · arXiv:2407.15545
-
LLMmap: Fingerprinting For Large Language Models 22 Jul 2024 · 1 repository · arXiv:2407.15847Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 16 harvested samples)
-
MMInstruct: A High-Quality Multi-Modal Instruction Tuning Dataset with Extensive Diversity 22 Jul 2024 · 1 repository · arXiv:2407.15838Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
MoRSE: Bridging the Gap in Cybersecurity Expertise with Retrieval Augmented Generation 22 Jul 2024 · 0 repositories · arXiv:2407.15748
-
RadioRAG: Factual large language models for enhanced diagnostics in radiology using online retrieval augmented generation 22 Jul 2024 · 1 repository · arXiv:2407.15621
-
Stretching Each Dollar: Diffusion Training from Scratch on a Micro-Budget 22 Jul 2024 · 1 repository · arXiv:2407.15811Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Unlocking the Potential: Benchmarking Large Language Models in Water Engineering and Research 22 Jul 2024 · 0 repositories · arXiv:2407.21045
-
ZZU-NLP at SIGHAN-2024 dimABSA Task: Aspect-Based Sentiment Analysis with Coarse-to-Fine In-context Learning 22 Jul 2024 · 0 repositories · arXiv:2407.15341
-
A multi-level multi-label text classification dataset of 19th century Ottoman and Russian literary and critical texts 21 Jul 2024 · 0 repositories · arXiv:2407.15136
-
Decoding Multilingual Moral Preferences: Unveiling LLM's Biases Through the Moral Machine Experiment 21 Jul 2024 · 1 repository · arXiv:2407.15184
-
Golden-Retriever: High-Fidelity Agentic Retrieval Augmented Generation for Industrial Knowledge Base 20 Jul 2024 · 0 repositories · arXiv:2408.00798
-
Automatic Generation of Fashion Images using Prompting in Generative Machine Learning Models 20 Jul 2024 · 1 repository · arXiv:2407.14944
-
Differential Privacy of Cross-Attention with Provable Guarantee 20 Jul 2024 · 0 repositories · arXiv:2407.14717
-
Adversarial Databases Improve Success in Retrieval-based Large Language Models 19 Jul 2024 · 0 repositories · arXiv:2407.14609
-
ChatQA 2: Bridging the Gap to Proprietary LLMs in Long Context and RAG Capabilities 19 Jul 2024 · 0 repositories · arXiv:2407.14482
-
Unipa-GPT: Large Language Models for university-oriented QA in Italian 19 Jul 2024 · 1 repository · arXiv:2407.14246
-
SQLfuse: Enhancing Text-to-SQL Performance through Comprehensive LLM Synergy 19 Jul 2024 · 0 repositories · arXiv:2407.14568
-
Scalable Exploration via Ensemble++ 18 Jul 2024 · 2 repositories · arXiv:2407.13195Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Black-Box Opinion Manipulation Attacks to Retrieval-Augmented Generation of Large Language Models 18 Jul 2024 · 0 repositories · arXiv:2407.13757
-
Can Open-Source LLMs Compete with Commercial Models? Exploring the Few-Shot Performance of Current GPT Models in Biomedical Tasks 18 Jul 2024 · 1 repository · arXiv:2407.13511
-
Evaluating Large Language Models for Anxiety and Depression Classification using Counseling and Psychotherapy Transcripts 18 Jul 2024 · 1 repository · arXiv:2407.13228
-
How Reliable are LLMs as Knowledge Bases? Re-thinking Facutality and Consistency 18 Jul 2024 · 0 repositories · arXiv:2407.13578
-
Learning-From-Mistakes Prompting for Indigenous Language Translation 18 Jul 2024 · 0 repositories · arXiv:2407.13343
-
PRAGyan -- Connecting the Dots in Tweets 18 Jul 2024 · 0 repositories · arXiv:2407.13909
-
Qalam : A Multimodal LLM for Arabic Optical Character and Handwriting Recognition 18 Jul 2024 · 0 repositories · arXiv:2407.13559
-
Reconfigurable Intelligent Surface Aided Vehicular Edge Computing: Joint Phase-shift Optimization and Multi-User Power Allocation 18 Jul 2024 · 1 repository · arXiv:2407.13123
-
Reconstruct the Pruned Model without Any Retraining 18 Jul 2024 · 0 repositories · arXiv:2407.13331
-
Retrieval-Augmented Generation for Natural Language Processing: A Survey 18 Jul 2024 · 0 repositories · arXiv:2407.13193
-
Retrieve, Summarize, Plan: Advancing Multi-hop Question Answering with an Iterative Approach 18 Jul 2024 · 0 repositories · arXiv:2407.13101
-
Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction 18 Jul 2024 · 1 repository · arXiv:2407.13943Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
AgentPoison: Red-teaming LLM Agents via Poisoning Memory or Knowledge Bases 17 Jul 2024 · 1 repository · arXiv:2407.12784Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 1 honoured, 0 violated, 14 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 19 harvested samples) · 1 pointer-only (licence)
-
Deep Learning-based Sentiment Analysis of Olympics Tweets 17 Jul 2024 · 0 repositories · arXiv:2407.12376
-
On Initializing Transformers with Pre-trained Embeddings 17 Jul 2024 · 0 repositories · arXiv:2407.12514
-
Optimizing Query Generation for Enhanced Document Retrieval in RAG 17 Jul 2024 · 0 repositories · arXiv:2407.12325
-
Evaluating Search Engines and Large Language Models for Answering Health Questions 17 Jul 2024 · 1 repository · arXiv:2407.12468
-
Sharif-STR at SemEval-2024 Task 1: Transformer as a Regression Model for Fine-Grained Scoring of Textual Semantic Relations 17 Jul 2024 · 1 repository · arXiv:2407.12426
-
Textualized and Feature-based Models for Compound Multimodal Emotion Recognition in the Wild 17 Jul 2024 · 2 repositories · arXiv:2407.12927Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Better RAG using Relevant Information Gain 16 Jul 2024 · 1 repository · arXiv:2407.12101
-
Beyond Binary: Multiclass Paraphasia Detection with Generative Pretrained Transformers and End-to-End Models 16 Jul 2024 · 0 repositories · arXiv:2407.11345
-
ChatBCG: Can AI Read Your Slide Deck? 16 Jul 2024 · 0 repositories · arXiv:2407.12875
-
Does Refusal Training in LLMs Generalize to the Past Tense? 16 Jul 2024 · 1 repository · arXiv:2407.11969Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
GPT Assisted Annotation of Rhetorical and Linguistic Features for Interpretable Propaganda Technique Detection in News Text 16 Jul 2024 · 0 repositories · arXiv:2407.11827
-
LaMI-DETR: Open-Vocabulary Detection with Language Model Instruction 16 Jul 2024 · 1 repository · arXiv:2407.11335Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Large Language Models as Misleading Assistants in Conversation 16 Jul 2024 · 0 repositories · arXiv:2407.11789
-
Large Visual-Language Models Are Also Good Classifiers: A Study of In-Context Multimodal Fake News Detection 16 Jul 2024 · 0 repositories · arXiv:2407.12879
-
LLMs-in-the-loop Part-1: Expert Small AI Models for Bio-Medical Text Translation 16 Jul 2024 · 0 repositories · arXiv:2407.12126
-
Mindful-RAG: A Study of Points of Failure in Retrieval Augmented Generation 16 Jul 2024 · 0 repositories · arXiv:2407.12216
-
R-SFLLM: Jamming Resilient Framework for Split Federated Learning with Large Language Models 16 Jul 2024 · 0 repositories · arXiv:2407.11654
-
Representation Bias in Political Sample Simulations with Large Language Models 16 Jul 2024 · 0 repositories · arXiv:2407.11409
-
ReFeR: Improving Evaluation and Reasoning through Hierarchy of Models 16 Jul 2024 · 0 repositories · arXiv:2407.12877
-
Scientific QA System with Verifiable Answers 16 Jul 2024 · 1 repository · arXiv:2407.11485
-
Trust No Bot: Discovering Personal Disclosures in Human-LLM Conversations in the Wild 16 Jul 2024 · 1 repository · arXiv:2407.11438
-
Communication- and Computation-Efficient Distributed Submodular Optimization in Robot Mesh Networks 15 Jul 2024 · 1 repository · arXiv:2407.10382
-
Deep Learning-Based Operators for Evolutionary Algorithms 15 Jul 2024 · 0 repositories · arXiv:2407.10477
-
CodeV: Empowering LLMs with HDL Generation through Multi-Level Summarization 15 Jul 2024 · 0 repositories · arXiv:2407.10424
-
Enhancing Retrieval and Managing Retrieval: A Four-Module Synergy for Improved Quality and Efficiency in RAG Systems 15 Jul 2024 · 1 repository · arXiv:2407.10670Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 1 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Evaluation of RAG Metrics for Question Answering in the Telecom Domain 15 Jul 2024 · 0 repositories · arXiv:2407.12873
-
Leveraging LLM-Respondents for Item Evaluation: a Psychometric Analysis 15 Jul 2024 · 0 repositories · arXiv:2407.10899
-
Making New Connections: LLMs as Puzzle Generators for The New York Times' Connections Word Game 15 Jul 2024 · 0 repositories · arXiv:2407.11240
-
Mechanistic interpretability of large language models with applications to the financial services industry 15 Jul 2024 · 0 repositories · arXiv:2407.11215
-
MetaLLM: A High-performant and Cost-efficient Dynamic Framework for Wrapping LLMs 15 Jul 2024 · 1 repository · arXiv:2407.10834Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
MixGR: Enhancing Retriever Generalization for Scientific Domain through Complementary Granularity 15 Jul 2024 · 1 repository · arXiv:2407.10691
-
Think-on-Graph 2.0: Deep and Faithful Large Language Model Reasoning with Knowledge-guided Retrieval Augmented Generation 15 Jul 2024 · 1 repository · arXiv:2407.10805Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 1 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 16 harvested samples)
-
Curriculum Learning for Small Code Language Models 14 Jul 2024 · 0 repositories · arXiv:2407.10194
-
Causality extraction from medical text using Large Language Models (LLMs) 13 Jul 2024 · 0 repositories · arXiv:2407.10020
-
Deep reinforcement learning with symmetric data augmentation applied for aircraft lateral attitude tracking control 13 Jul 2024 · 0 repositories · arXiv:2407.11077
-
Document-level Clinical Entity and Relation Extraction via Knowledge Base-Guided Generation 13 Jul 2024 · 0 repositories · arXiv:2407.10021
-
Generating In-store Customer Journeys from Scratch with GPT Architectures 13 Jul 2024 · 0 repositories · arXiv:2407.11081
-
Hydra: Bidirectional State Space Models Through Generalized Matrix Mixers 13 Jul 2024 · 1 repository · arXiv:2407.09941Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Resource Management for Low-latency Cooperative Fine-tuning of Foundation Models at the Network Edge 13 Jul 2024 · 0 repositories · arXiv:2407.09873
-
ASTPrompter: Weakly Supervised Automated Language Model Red-Teaming to Identify Low-Perplexity Toxic Prompts 12 Jul 2024 · 1 repository · arXiv:2407.09447