Methods › General › Regularization › Weight Decay › Papers, page 46
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 46 of 108: papers 4,501 to 4,600 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
TextrolSpeech: A Text Style Control Speech Corpus With Codec Language Text-to-Speech Models 28 Aug 2023 · 1 repository · arXiv:2308.14430
-
Examining User-Friendly and Open-Sourced Large GPT Models: A Survey on Language, Multimodal, and Scientific GPT Models 27 Aug 2023 · 1 repository · arXiv:2308.14149
-
A Wide Evaluation of ChatGPT on Affective Computing Tasks 26 Aug 2023 · 1 repository · arXiv:2308.13911
-
Improving Knowledge Distillation for BERT Models: Loss Functions, Mapping Methods, and Weight Tuning 26 Aug 2023 · 0 repositories · arXiv:2308.13958
-
Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions 25 Aug 2023 · 1 repository · arXiv:2309.12342
-
Leveraging Knowledge and Reinforcement Learning for Enhanced Reliability of Language Models 25 Aug 2023 · 0 repositories · arXiv:2308.13467
-
MLLM-DataEngine: An Iterative Refinement Approach for MLLM 25 Aug 2023 · 1 repository · arXiv:2308.13566Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 1 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
Transforming the Output of Generative Pre-trained Transformer: The Influence of the PGI Framework on Attention Dynamics 25 Aug 2023 · 0 repositories · arXiv:2308.13317
-
A Small and Fast BERT for Chinese Medical Punctuation Restoration 24 Aug 2023 · 1 repository · arXiv:2308.12568
-
Advancing Hungarian Text Processing with HuSpaCy: Efficient and Accurate NLP Pipelines 24 Aug 2023 · 2 repositories · arXiv:2308.12635
-
Financial News Analytics Using Fine-Tuned Llama 2 GPT Model 24 Aug 2023 · 0 repositories · arXiv:2308.13032
-
Multi-BERT for Embeddings for Recommendation System 24 Aug 2023 · 0 repositories · arXiv:2308.13050
-
Sentence Embedding Models for Ancient Greek Using Multilingual Knowledge Distillation 24 Aug 2023 · 2 repositories · arXiv:2308.13116
-
Text Similarity from Image Contents using Statistical and Semantic Analysis Techniques 24 Aug 2023 · 0 repositories · arXiv:2308.12842
-
Simple is Better and Large is Not Enough: Towards Ensembling of Foundational Language Models 23 Aug 2023 · 0 repositories · arXiv:2308.12272
-
Evaluating Large Language Models on Graphs: Performance Insights and Comparative Analysis 22 Aug 2023 · 1 repository · arXiv:2308.11224Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Exploring the Effectiveness of GPT Models in Test-Taking: A Case Study of the Driver's License Knowledge Test 22 Aug 2023 · 0 repositories · arXiv:2308.11827
-
MulMarker: a comprehensive framework for identifying multi-gene prognostic signatures 22 Aug 2023 · 1 repository · arXiv:2308.11349
-
Tryage: Real-time, intelligent Routing of User Prompts to Large Language Models 22 Aug 2023 · 0 repositories · arXiv:2308.11601
-
GPT-in-the-Loop: Adaptive Decision-Making for Multiagent Systems 21 Aug 2023 · 0 repositories · arXiv:2308.10435
-
GradientCoin: A Peer-to-Peer Decentralized Large Language Models 21 Aug 2023 · 0 repositories · arXiv:2308.10502
-
PlatoLM: Teaching LLMs in Multi-Round Dialogue via a User Simulator 21 Aug 2023 · 1 repository · arXiv:2308.11534Syntology 0 ran · 2 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Large Language Models on Wikipedia-Style Survey Generation: an Evaluation in NLP Concepts 21 Aug 2023 · 1 repository · arXiv:2308.10410
-
SpikingBERT: Distilling BERT to Train Spiking Language Models Using Implicit Differentiation 21 Aug 2023 · 1 repository · arXiv:2308.10873Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 2 honoured, 0 violated, 6 with no contract checked; 5 where Syntology's instrument failed) · 4 unverified (of 17 harvested samples) · 17 pointer-only (licence)
-
Steering Language Models With Activation Engineering 20 Aug 2023 · 2 repositories · arXiv:2308.10248Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
How Good Are LLMs at Out-of-Distribution Detection? 20 Aug 2023 · 1 repository · arXiv:2308.10261Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Improving Adversarial Robustness of Masked Autoencoders via Test-time Frequency-domain Prompting 20 Aug 2023 · 1 repository · arXiv:2308.10315
-
Data-to-text Generation for Severely Under-Resourced Languages with GPT-3.5: A Bit of Help Needed from Google Translate 19 Aug 2023 · 1 repository · arXiv:2308.09957
-
East: Efficient and Accurate Secure Transformer Framework for Inference 19 Aug 2023 · 0 repositories · arXiv:2308.09923
-
Open, Closed, or Small Language Models for Text Classification? 19 Aug 2023 · 0 repositories · arXiv:2308.10092
-
Optimizing Multi-Class Text Classification: A Diverse Stacking Ensemble Framework Utilizing Transformers 19 Aug 2023 · 0 repositories · arXiv:2308.11519
-
A tailored Handwritten-Text-Recognition System for Medieval Latin 18 Aug 2023 · 0 repositories · arXiv:2308.09368
-
How susceptible are LLMs to Logical Fallacies? 18 Aug 2023 · 1 repository · arXiv:2308.09853
-
Learning Representations on Logs for AIOps 18 Aug 2023 · 1 repository · arXiv:2308.11526
-
Predictive Authoring for Brazilian Portuguese Augmentative and Alternative Communication 18 Aug 2023 · 1 repository · arXiv:2308.09497
-
WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct 18 Aug 2023 · 1 repository · arXiv:2308.09583Syntology 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 6 unverified (of 16 harvested samples) · 16 pointer-only (licence)
-
A Comparative Study of Text Embedding Models for Semantic Text Similarity in Bug Reports 17 Aug 2023 · 1 repository · arXiv:2308.09193
-
End-to-End Beam Retrieval for Multi-Hop Question Answering 17 Aug 2023 · 3 repositories · arXiv:2308.08973Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 2 pointer-only (licence)
-
Evaluation of really good grammatical error correction 17 Aug 2023 · 1 repository · arXiv:2308.08982
-
MaScQA: A Question Answering Dataset for Investigating Materials Science Knowledge of Large Language Models 17 Aug 2023 · 0 repositories · arXiv:2308.09115
-
MindMap: Knowledge Graph Prompting Sparks Graph of Thoughts in Large Language Models 17 Aug 2023 · 1 repository · arXiv:2308.09729Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
A Preliminary Study on a Conceptual Game Feature Generation and Recommendation System 16 Aug 2023 · 0 repositories · arXiv:2308.13538
-
BIOptimus: Pre-training an Optimal Biomedical Language Model with Curriculum Learning for Named Entity Recognition 16 Aug 2023 · 1 repository · arXiv:2308.08625
-
Self-Deception: Reverse Penetrating the Semantic Firewall of Large Language Models 16 Aug 2023 · 0 repositories · arXiv:2308.11521
-
Forward-Backward Reasoning in Large Language Models for Mathematical Verification 15 Aug 2023 · 0 repositories · arXiv:2308.07758
-
"Beware of deception": Detecting Half-Truth and Debunking it through Controlled Claim Editing 15 Aug 2023 · 0 repositories · arXiv:2308.07973
-
CALYPSO: LLMs as Dungeon Masters' Assistants 15 Aug 2023 · 0 repositories · arXiv:2308.07540
-
DS4DH at #SMM4H 2023: Zero-Shot Adverse Drug Events Normalization using Sentence Transformers and Reciprocal-Rank Fusion 15 Aug 2023 · 0 repositories · arXiv:2308.12877
-
Finding Stakeholder-Material Information from 10-K Reports using Fine-Tuned BERT and LSTM Models 15 Aug 2023 · 0 repositories · arXiv:2308.07522
-
From Commit Message Generation to History-Aware Commit Message Completion 15 Aug 2023 · 1 repository · arXiv:2308.07655
-
MultiSChuBERT: Effective Multimodal Fusion for Scholarly Document Quality Prediction 15 Aug 2023 · 0 repositories · arXiv:2308.07971
-
SPM: Structured Pretraining and Matching Architectures for Relevance Modeling in Meituan Search 15 Aug 2023 · 0 repositories · arXiv:2308.07711
-
Leveraging Codebook Knowledge with NLI and ChatGPT for Zero-Shot Political Relation Classification 15 Aug 2023 · 1 repository · arXiv:2308.07876
-
Ternary Singular Value Decomposition as a Better Parameterized Form in Linear Mapping 15 Aug 2023 · 1 repository · arXiv:2308.07641Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Approximating Human-Like Few-shot Learning with GPT-based Compression 14 Aug 2023 · 0 repositories · arXiv:2308.06942
-
Generating Individual Trajectories Using GPT-2 Trained from Scratch on Encoded Spatiotemporal Data 14 Aug 2023 · 0 repositories · arXiv:2308.07940
-
LLM Self Defense: By Self Examination, LLMs Know They Are Being Tricked 14 Aug 2023 · 1 repository · arXiv:2308.07308
-
Playing with Words: Comparing the Vocabulary and Lexical Richness of ChatGPT and Humans 14 Aug 2023 · 0 repositories · arXiv:2308.07462
-
Semantic Similarity Loss for Neural Source Code Summarization 14 Aug 2023 · 1 repository · arXiv:2308.07429
-
An Ensemble Approach to Question Classification: Integrating Electra Transformer, GloVe, and LSTM 13 Aug 2023 · 0 repositories · arXiv:2308.06828
-
Improving Face Recognition from Caption Supervision with Multi-Granular Contextual Feature Aggregation 13 Aug 2023 · 0 repositories · arXiv:2308.06866
-
Assessing Student Errors in Experimentation Using Artificial Intelligence and Large Language Models: A Comparative Study with Human Raters 11 Aug 2023 · 0 repositories · arXiv:2308.06088
-
Enhancing Phenotype Recognition in Clinical Notes Using Large Language Models: PhenoBCBERT and PhenoGPT 11 Aug 2023 · 1 repository · arXiv:2308.06294
-
Identification of the Relevance of Comments in Codes Using Bag of Words and Transformer Based Models 11 Aug 2023 · 1 repository · arXiv:2308.06144
-
Large Language Models in Cryptocurrency Securities Cases: Can a GPT Model Meaningfully Assist Lawyers? 11 Aug 2023 · 0 repositories · arXiv:2308.06032
-
Large Language Models to Identify Social Determinants of Health in Electronic Health Records 11 Aug 2023 · 1 repository · arXiv:2308.06354Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Task Conditioned BERT for Joint Intent Detection and Slot-filling 11 Aug 2023 · 0 repositories · arXiv:2308.06165
-
Adaptive Low Rank Adaptation of Segment Anything to Salient Object Detection 10 Aug 2023 · 1 repository · arXiv:2308.05426
-
AudioLDM 2: Learning Holistic Audio Generation with Self-supervised Pretraining 10 Aug 2023 · 2 repositories · arXiv:2308.05734Syntology official (archive's flag): 8 ran · 16 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 2 honoured, 3 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 11 unverified (of 27 harvested samples) · 19 pointer-only (licence)
-
Bringing order into the realm of Transformer-based language models for artificial intelligence and law 10 Aug 2023 · 0 repositories · arXiv:2308.05502
-
Exploring Machine Learning and Transformer-based Approaches for Deceptive Text Classification: A Comparative Analysis 10 Aug 2023 · 0 repositories · arXiv:2308.05476
-
Metacognitive Prompting Improves Understanding in Large Language Models 10 Aug 2023 · 1 repository · arXiv:2308.05342
-
RTLLM: An Open-Source Benchmark for Design RTL Generation with Large Language Model 10 Aug 2023 · 1 repository · arXiv:2308.05345Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Testing GPT-4 with Wolfram Alpha and Code Interpreter plug-ins on math and science problems 10 Aug 2023 · 0 repositories · arXiv:2308.05713
-
WeaverBird: Empowering Financial Decision-Making with Large Language Model, Knowledge Base, and Search Engine 10 Aug 2023 · 1 repository · arXiv:2308.05361
-
You Only Prompt Once: On the Capabilities of Prompt Learning on Large Language Models to Tackle Toxic Content 10 Aug 2023 · 1 repository · arXiv:2308.05596Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
An Empirical Study on Using Large Language Models to Analyze Software Supply Chain Security Failures 9 Aug 2023 · 0 repositories · arXiv:2308.04898
-
LLaMA-E: Empowering E-commerce Authoring with Object-Interleaved Instruction Following 9 Aug 2023 · 0 repositories · arXiv:2308.04913
-
LLMeBench: A Flexible Framework for Accelerating LLMs Benchmarking 9 Aug 2023 · 1 repository · arXiv:2308.04945
-
MetRoBERTa: Leveraging Traditional Customer Relationship Management Data to Develop a Transit-Topic-Aware Language Model 9 Aug 2023 · 0 repositories · arXiv:2308.05012
-
Performance Analysis of Transformer Based Models (BERT, ALBERT and RoBERTa) in Fake News Detection 9 Aug 2023 · 1 repository · arXiv:2308.04950
-
3D-VisTA: Pre-trained Transformer for 3D Vision and Text Alignment 8 Aug 2023 · 1 repository · arXiv:2308.04352Syntology 4 ran (of which 2 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 1 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Gromov-Wasserstein unsupervised alignment reveals structural correspondences between the color similarity structures of humans and large language models 8 Aug 2023 · 0 repositories · arXiv:2308.04381
-
I-WAS: a Data Augmentation Method with GPT-2 for Simile Detection 8 Aug 2023 · 0 repositories · arXiv:2308.04109
-
A Cross-Domain Evaluation of Approaches for Causal Knowledge Extraction 7 Aug 2023 · 1 repository · arXiv:2308.03891
-
Analysis of the Evolution of Advanced Transformer-Based Language Models: Experiments on Opinion Mining 7 Aug 2023 · 1 repository · arXiv:2308.03235
-
Detecting Spells in Fantasy Literature with a Transformer Based Artificial Intelligence 7 Aug 2023 · 0 repositories · arXiv:2308.03660
-
Fact-Checking Generative AI: Ontology-Driven Biological Graphs for Disease-Gene Link Verification 7 Aug 2023 · 0 repositories · arXiv:2308.03929
-
Exploring ChatGPT's Empathic Abilities 7 Aug 2023 · 1 repository · arXiv:2308.03527
-
CORAL: Expert-Curated medical Oncology Reports to Advance Language Model Inference 7 Aug 2023 · 1 repository · arXiv:2308.03853
-
KITLM: Domain-Specific Knowledge InTegration into Language Models for Question Answering 7 Aug 2023 · 1 repository · arXiv:2308.03638
-
Topological Interpretations of GPT-3 7 Aug 2023 · 0 repositories · arXiv:2308.03565
-
Trusting Language Models in Education 7 Aug 2023 · 0 repositories · arXiv:2308.03866
-
Training BERT Models to Carry Over a Coding System Developed on One Corpus to Another 7 Aug 2023 · 0 repositories · arXiv:2308.03742
-
GPTScan: Detecting Logic Vulnerabilities in Smart Contracts by Combining GPT with Program Analysis 7 Aug 2023 · 1 repository · arXiv:2308.03314
-
End-to-End Query Term Weighting 6 Aug 2023 · 1 repository
-
"Kurosawa": A Script Writer's Assistant 6 Aug 2023 · 0 repositories · arXiv:2308.03122
-
TARJAMAT: Evaluation of Bard and ChatGPT on Machine Translation of Ten Arabic Varieties 6 Aug 2023 · 0 repositories · arXiv:2308.03051
-
DaMSTF: Domain Adversarial Learning Enhanced Meta Self-Training for Domain Adaptation 5 Aug 2023 · 0 repositories · arXiv:2308.02753Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples)
-
An APT Event Extraction Method Based on BERT-BiGRU-CRF for APT Attack Detection 4 Aug 2023 · 0 repositories