Methods › General › Regularization › Weight Decay › Papers, page 45
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 45 of 108: papers 4,401 to 4,500 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
EvoPrompt: Connecting LLMs with Evolutionary Algorithms Yields Powerful Prompt Optimizers 15 Sep 2023 · 2 repositories · arXiv:2309.08532Syntology 11 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 8 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples)
-
CoCA: Fusing Position Embedding with Collinear Constrained Attention in Transformers for Long Context Window Extending 15 Sep 2023 · 1 repository · arXiv:2309.08646
-
Detecting Relevant Information in High-Volume Chat Logs: Keyphrase Extraction for Grooming and Drug Dealing Forensic Analysis 15 Sep 2023 · 0 repositories · arXiv:2311.04905
-
GPT-Lab: Next Generation Of Optimal Chemistry Discovery By GPT Driven Robotic Lab 15 Sep 2023 · 0 repositories · arXiv:2309.16721
-
ICLEF: In-Context Learning with Expert Feedback for Explainable Style Transfer 15 Sep 2023 · 1 repository · arXiv:2309.08583
-
Large Language Models for Failure Mode Classification: An Investigation 15 Sep 2023 · 1 repository · arXiv:2309.08181
-
Structural Self-Supervised Objectives for Transformers 15 Sep 2023 · 1 repository · arXiv:2309.08272
-
Transformer Based Punctuation Restoration for Turkish 15 Sep 2023 · 1 repository
-
VulnSense: Efficient Vulnerability Detection in Ethereum Smart Contracts by Multimodal Learning with Graph Neural Network and Language Model 15 Sep 2023 · 0 repositories · arXiv:2309.08474
-
An Empirical Evaluation of Prompting Strategies for Large Language Models in Zero-Shot Clinical Natural Language Processing 14 Sep 2023 · 0 repositories · arXiv:2309.08008
-
Assessing the nature of large language models: A caution against anthropocentrism 14 Sep 2023 · 0 repositories · arXiv:2309.07683
-
Automatic Data Visualization Generation from Chinese Natural Language Questions 14 Sep 2023 · 0 repositories · arXiv:2309.07650
-
ChatGPT MT: Competitive for High- (but not Low-) Resource Languages 14 Sep 2023 · 2 repositories · arXiv:2309.07423Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
DebCSE: Rethinking Unsupervised Contrastive Sentence Embedding Learning in the Debiasing Perspective 14 Sep 2023 · 0 repositories · arXiv:2309.07396
-
EnCodecMAE: Leveraging neural codecs for universal audio representation learning 14 Sep 2023 · 2 repositories · arXiv:2309.07391Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Text Classification of Cancer Clinical Trial Eligibility Criteria 14 Sep 2023 · 0 repositories · arXiv:2309.07812
-
Two Timin': Repairing Smart Contracts With A Two-Layered Approach 14 Sep 2023 · 0 repositories · arXiv:2309.07841
-
Large Language Models Can Infer Psychological Dispositions of Social Media Users 13 Sep 2023 · 0 repositories · arXiv:2309.08631
-
Traveling Words: A Geometric Interpretation of Transformers 13 Sep 2023 · 1 repository · arXiv:2309.07315
-
Balanced and Explainable Social Media Analysis for Public Health with Large Language Models 12 Sep 2023 · 1 repository · arXiv:2309.05951
-
Circuit Breaking: Removing Model Behaviors with Targeted Ablation 12 Sep 2023 · 1 repository · arXiv:2309.05973Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Characterizing Latent Perspectives of Media Houses Towards Public Figures 12 Sep 2023 · 0 repositories · arXiv:2309.06112
-
PRESTI: Predicting Repayment Effort of Self-Admitted Technical Debt Using Textual Information 12 Sep 2023 · 0 repositories · arXiv:2309.06020
-
Comparing Llama-2 and GPT-3 LLMs for HPC kernels generation 12 Sep 2023 · 0 repositories · arXiv:2309.07103
-
Exploring Large Language Models for Ontology Alignment 12 Sep 2023 · 1 repository · arXiv:2309.07172
-
Overview of Memotion 3: Sentiment and Emotion Analysis of Codemixed Hinglish Memes 12 Sep 2023 · 0 repositories · arXiv:2309.06517
-
Strategic Behavior of Large Language Models: Game Structure vs. Contextual Framing 12 Sep 2023 · 0 repositories · arXiv:2309.05898
-
The Moral Machine Experiment on Large Language Models 12 Sep 2023 · 1 repository · arXiv:2309.05958
-
Unveiling the potential of large language models in generating semantic and cross-language clones 12 Sep 2023 · 0 repositories · arXiv:2309.06424
-
Applying BioBERT to Extract Germline Gene-Disease Associations for Building a Knowledge Graph from the Biomedical Literature 11 Sep 2023 · 1 repository · arXiv:2309.13061
-
Black-Box Analysis: GPTs Across Time in Legal Textual Entailment Task 11 Sep 2023 · 0 repositories · arXiv:2309.05501
-
CrisisTransformers: Pre-trained language models and sentence encoders for crisis-related social media texts 11 Sep 2023 · 0 repositories · arXiv:2309.05494
-
Detecting Natural Language Biases with Prompt-based Learning 11 Sep 2023 · 0 repositories · arXiv:2309.05227
-
Memory Injections: Correcting Multi-Hop Reasoning Failures during Inference in Transformer-Based Language Models 11 Sep 2023 · 1 repository · arXiv:2309.05605Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
SparseSwin: Swin Transformer with Sparse Transformer Block 11 Sep 2023 · 1 repository · arXiv:2309.05224
-
Zero-shot Learning with Minimum Instruction to Extract Social Determinants and Family History from Clinical Notes using GPT Model 11 Sep 2023 · 0 repositories · arXiv:2309.05475
-
Implementing Learning Principles with a Personal AI Tutor: A Case Study 10 Sep 2023 · 0 repositories · arXiv:2309.13060
-
Learning Personalized User Preference from Cold Start in Multi-turn Conversations 10 Sep 2023 · 0 repositories · arXiv:2309.05127
-
Neural-Hidden-CRF: A Robust Weakly-Supervised Sequence Labeler 10 Sep 2023 · 1 repository · arXiv:2309.05086
-
RGAT: A Deeper Look into Syntactic Dependency Information for Coreference Resolution 10 Sep 2023 · 0 repositories · arXiv:2309.04977
-
Towards Understanding Neural Collapse: The Effects of Batch Normalization and Weight Decay 9 Sep 2023 · 0 repositories · arXiv:2309.04644
-
Can NLP Models 'Identify', 'Distinguish', and 'Justify' Questions that Don't have a Definitive Answer? 8 Sep 2023 · 0 repositories · arXiv:2309.04635
-
Context-Aware Prompt Tuning for Vision-Language Model with Dual-Alignment 8 Sep 2023 · 0 repositories · arXiv:2309.04158
-
Encoding Multi-Domain Scientific Papers by Ensembling Multiple CLS Tokens 8 Sep 2023 · 1 repository · arXiv:2309.04333Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Fuzzy Fingerprinting Transformer Language-Models for Emotion Recognition in Conversations 8 Sep 2023 · 0 repositories · arXiv:2309.04292
-
Leveraging Pretrained Image-text Models for Improving Audio-Visual Learning 8 Sep 2023 · 0 repositories · arXiv:2309.04628
-
UQ at #SMM4H 2023: ALEX for Public Health Analysis with Social Media 8 Sep 2023 · 1 repository · arXiv:2309.04213
-
Evaluating ChatGPT as a Recommender System: A Rigorous Approach 7 Sep 2023 · 1 repository · arXiv:2309.03613
-
Supervised Learning and Large Language Model Benchmarks on Mental Health Datasets: Cognitive Distortions and Suicidal Risks in Chinese Social Media 7 Sep 2023 · 2 repositories · arXiv:2309.03564
-
FLM-101B: An Open LLM and How to Train It with $100K Budget 7 Sep 2023 · 0 repositories · arXiv:2309.03852
-
Zero-Shot Audio Captioning via Audibility Guidance 7 Sep 2023 · 0 repositories · arXiv:2309.03884
-
Certifying LLM Safety against Adversarial Prompting 6 Sep 2023 · 1 repository · arXiv:2309.02705
-
HAE-RAE Bench: Evaluation of Korean Knowledge in Language Models 6 Sep 2023 · 1 repository · arXiv:2309.02706
-
Leave no Place Behind: Improved Geolocation in Humanitarian Documents 6 Sep 2023 · 0 repositories · arXiv:2309.02914
-
Offensive Hebrew Corpus and Detection using BERT 6 Sep 2023 · 1 repository · arXiv:2309.02724
-
Self-Supervised Masked Digital Elevation Models Encoding for Low-Resource Downstream Tasks 6 Sep 2023 · 0 repositories · arXiv:2309.03367
-
CodeApex: A Bilingual Programming Evaluation Benchmark for Large Language Models 5 Sep 2023 · 1 repository · arXiv:2309.01940
-
Do You Trust ChatGPT? -- Perceived Credibility of Human and AI-Generated Content 5 Sep 2023 · 0 repositories · arXiv:2309.02524
-
Incorporating Dictionaries into a Neural Network Architecture to Extract COVID-19 Medical Concepts From Social Media 5 Sep 2023 · 0 repositories · arXiv:2309.02188
-
Language Models for Novelty Detection in System Call Traces 5 Sep 2023 · 0 repositories · arXiv:2309.02206
-
Leveraging BERT Language Models for Multi-Lingual ESG Issue Identification 5 Sep 2023 · 0 repositories · arXiv:2309.02189
-
Sample Size in Natural Language Processing within Healthcare Research 5 Sep 2023 · 0 repositories · arXiv:2309.02237
-
Benchmarking Large Language Models in Retrieval-Augmented Generation 4 Sep 2023 · 1 repository · arXiv:2309.01431Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Do androids dream of fictional references? A bibliographic dialogue with ChatGPT3.5 4 Sep 2023 · 0 repositories · arXiv:2312.00789
-
No Data Augmentation? Alternative Regularizations for Effective Training on Small Datasets 4 Sep 2023 · 0 repositories · arXiv:2309.01694
-
Prompting or Fine-tuning? A Comparative Study of Large Language Models for Taxonomy Construction 4 Sep 2023 · 1 repository · arXiv:2309.01715
-
A Study on the Implementation of Generative AI Services Using an Enterprise Data-Based LLM Application Architecture 3 Sep 2023 · 0 repositories · arXiv:2309.01105
-
A Visual Interpretation-Based Self-Improved Classification System Using Virtual Adversarial Training 3 Sep 2023 · 0 repositories · arXiv:2309.01196
-
Saturn: An Optimized Data System for Large Model Deep Learning Workloads 3 Sep 2023 · 1 repository · arXiv:2309.01226
-
Knowledge Graph Embeddings for Multi-Lingual Structured Representations of Radiology Reports 2 Sep 2023 · 0 repositories · arXiv:2309.00917
-
Studying the impacts of pre-training using ChatGPT-generated text on downstream tasks 2 Sep 2023 · 0 repositories · arXiv:2309.05668
-
BatchPrompt: Accomplish more with less 1 Sep 2023 · 1 repository · arXiv:2309.00384
-
Large Language Models for Semantic Monitoring of Corporate Disclosures: A Case Study on Korea's Top 50 KOSPI Companies 1 Sep 2023 · 0 repositories · arXiv:2309.00208
-
Publicly Shareable Clinical Large Language Model Built on Synthetic Clinical Notes 1 Sep 2023 · 1 repository · arXiv:2309.00237Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
SortedNet: A Scalable and Generalized Framework for Training Modular Deep Neural Networks 1 Sep 2023 · 0 repositories · arXiv:2309.00255
-
Taken out of context: On measuring situational awareness in LLMs 1 Sep 2023 · 1 repository · arXiv:2309.00667Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Why do universal adversarial attacks work on large language models?: Geometry might be the answer 1 Sep 2023 · 0 repositories · arXiv:2309.00254
-
BioCoder: A Benchmark for Bioinformatics Code Generation with Large Language Models 31 Aug 2023 · 1 repository · arXiv:2308.16458Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Can humans help BERT gain "confidence"? 31 Aug 2023 · 0 repositories · arXiv:2309.06580
-
DictaBERT: A State-of-the-Art BERT Suite for Modern Hebrew 31 Aug 2023 · 0 repositories · arXiv:2308.16687
-
GPT has become financially literate: Insights from financial literacy tests of GPT and a preliminary test of how people use it as a source of advice 31 Aug 2023 · 0 repositories · arXiv:2309.00649
-
Ladder-of-Thought: Using Knowledge as Steps to Elevate Stance Detection 31 Aug 2023 · 0 repositories · arXiv:2308.16763
-
Linking microblogging sentiments to stock price movement: An application of GPT-4 31 Aug 2023 · 0 repositories · arXiv:2308.16771
-
SARATHI: Efficient LLM Inference by Piggybacking Decodes with Chunked Prefills 31 Aug 2023 · 0 repositories · arXiv:2308.16369
-
Towards Improving the Expressiveness of Singing Voice Synthesis with BERT Derived Semantic Information 31 Aug 2023 · 0 repositories · arXiv:2308.16836
-
SP³: Enhancing Structured Pruning via PCA Projection 31 Aug 2023 · 1 repository · arXiv:2308.16475Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 1 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples) · 14 pointer-only (licence)
-
Analyzing Character and Consciousness in AI-Generated Social Content: A Case Study of Chirper, the AI Social Network 30 Aug 2023 · 0 repositories · arXiv:2309.08614
-
Jais and Jais-chat: Arabic-Centric Foundation and Instruction-Tuned Open Generative Large Language Models 30 Aug 2023 · 0 repositories · arXiv:2308.16149
-
Large Language Models as Data Preprocessors 30 Aug 2023 · 0 repositories · arXiv:2308.16361
-
Quantifying Uncertainty in Answers from any Language Model and Enhancing their Trustworthiness 30 Aug 2023 · 0 repositories · arXiv:2308.16175
-
Response: Emergent analogical reasoning in large language models 30 Aug 2023 · 1 repository · arXiv:2308.16118
-
The DeepZen Speech Synthesis System for Blizzard Challenge 2023 30 Aug 2023 · 0 repositories · arXiv:2308.15945
-
FurChat: An Embodied Conversational Agent using LLMs, Combining Open and Closed-Domain Dialogue with Facial Expressions 29 Aug 2023 · 0 repositories · arXiv:2308.15214
-
Multi-party Goal Tracking with LLMs: Comparing Pre-training, Fine-tuning, and Prompt Engineering 29 Aug 2023 · 1 repository · arXiv:2308.15231
-
SpikeBERT: A Language Spikformer Learned from BERT with Knowledge Distillation 29 Aug 2023 · 1 repository · arXiv:2308.15122Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
ANER: Arabic and Arabizi Named Entity Recognition using Transformer-Based Approach 28 Aug 2023 · 1 repository · arXiv:2308.14669
-
Breaking the Bank with ChatGPT: Few-Shot Text Classification for Finance 28 Aug 2023 · 0 repositories · arXiv:2308.14634
-
Cognitive Effects in Large Language Models 28 Aug 2023 · 1 repository · arXiv:2308.14337
-
Distilled GPT for Source Code Summarization 28 Aug 2023 · 1 repository · arXiv:2308.14731
-
Target-independent XLA optimization using Reinforcement Learning 28 Aug 2023 · 0 repositories · arXiv:2308.14364