Methods › General › Regularization › Weight Decay › Papers, page 52
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 52 of 108: papers 5,101 to 5,200 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Can NLP Models Correctly Reason Over Contexts that Break the Common Assumptions? 20 May 2023 · 0 repositories · arXiv:2305.12096
-
CDJUR-BR -- A Golden Collection of Legal Document from Brazilian Justice with Fine-Grained Named Entities 20 May 2023 · 0 repositories · arXiv:2305.18315
-
LogiCoT: Logical Chain-of-Thought Instruction-Tuning 20 May 2023 · 1 repository · arXiv:2305.12147
-
Practical PCG Through Large Language Models 20 May 2023 · 0 repositories · arXiv:2305.18243
-
Revisiting the Architectures like Pointer Networks to Efficiently Improve the Next Word Distribution, Summarization Factuality, and Beyond 20 May 2023 · 1 repository · arXiv:2305.12289
-
SEntFiN 1.0: Entity-Aware Sentiment Analysis for Financial News 20 May 2023 · 0 repositories · arXiv:2305.12257
-
A Sequence-to-Sequence Approach for Arabic Pronoun Resolution 19 May 2023 · 0 repositories · arXiv:2305.11529
-
AutoTrial: Prompting Language Models for Clinical Trial Design 19 May 2023 · 0 repositories · arXiv:2305.11366
-
Clinical Camel: An Open Expert-Level Medical Language Model with Dialogue-Based Knowledge Encoding 19 May 2023 · 2 repositories · arXiv:2305.12031Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Exploring the Upper Limits of Text-Based Collaborative Filtering Using Large Language Models: Discoveries and Insights 19 May 2023 · 0 repositories · arXiv:2305.11700
-
Eye-SpatialNet: Spatial Information Extraction from Ophthalmology Notes 19 May 2023 · 0 repositories · arXiv:2305.11948
-
Federated Foundation Models: Privacy-Preserving and Collaborative Learning for Large Models 19 May 2023 · 0 repositories · arXiv:2305.11414
-
Language-Universal Phonetic Representation in Multilingual Speech Pretraining for Low-Resource Speech Recognition 19 May 2023 · 0 repositories · arXiv:2305.11569
-
PointGPT: Auto-regressively Generative Pre-training from Point Clouds 19 May 2023 · 1 repository · arXiv:2305.11487Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 4 pointer-only (licence)
-
Scaling laws for language encoding models in fMRI 19 May 2023 · 1 repository · arXiv:2305.11863
-
SeeGULL: A Stereotype Benchmark with Broad Geo-Cultural Coverage Leveraging Generative Models 19 May 2023 · 1 repository · arXiv:2305.11840
-
Self-Agreement: A Framework for Fine-tuning Language Models to Find Agreement among Diverse Opinions 19 May 2023 · 0 repositories · arXiv:2305.11460
-
Hint of Thought prompting: an explainable and zero-shot approach to reasoning tasks with LLMs 19 May 2023 · 0 repositories · arXiv:2305.11461
-
Ahead-of-Time P-Tuning 18 May 2023 · 0 repositories · arXiv:2305.10835
-
AIwriting: Relations Between Image Generation and Digital Writing 18 May 2023 · 0 repositories · arXiv:2305.10834
-
Comparing Machines and Children: Using Developmental Psychology Experiments to Assess the Strengths and Weaknesses of LaMDA Responses 18 May 2023 · 0 repositories · arXiv:2305.11243
-
Deep Learning Methods for Extracting Metaphorical Names of Flowers and Plants 18 May 2023 · 0 repositories · arXiv:2305.10833
-
Ditto: A Simple and Efficient Approach to Improve Sentence Embeddings 18 May 2023 · 1 repository · arXiv:2305.10786
-
Generalized Multiple Intent Conditioned Slot Filling 18 May 2023 · 0 repositories · arXiv:2305.11023
-
Generalized Planning in PDDL Domains with Pretrained Large Language Models 18 May 2023 · 1 repository · arXiv:2305.11014Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Large Language Models can be Guided to Evade AI-Generated Text Detection 18 May 2023 · 1 repository · arXiv:2305.10847
-
PDP: Parameter-free Differentiable Pruning is All You Need 18 May 2023 · 0 repositories · arXiv:2305.11203
-
Trading Syntax Trees for Wordpieces: Target-oriented Opinion Words Extraction with Wordpieces and Aspect Enhancement 18 May 2023 · 0 repositories · arXiv:2305.11034
-
A quantitative study of NLP approaches to question difficulty estimation 17 May 2023 · 1 repository · arXiv:2305.10236
-
AD-KD: Attribution-Driven Knowledge Distillation for Language Model Compression 17 May 2023 · 1 repository · arXiv:2305.10010Syntology official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
CoEdIT: Text Editing by Task-Specific Instruction Tuning 17 May 2023 · 1 repository · arXiv:2305.09857
-
Explaining black box text modules in natural language with language models 17 May 2023 · 2 repositories · arXiv:2305.09863Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
From chocolate bunny to chocolate crocodile: Do Language Models Understand Noun Compounds? 17 May 2023 · 0 repositories · arXiv:2305.10568
-
Interactive Learning of Hierarchical Tasks from Dialog with GPT 17 May 2023 · 0 repositories · arXiv:2305.10349
-
Knowledge Graph Completion Models are Few-shot Learners: An Empirical Study of Relation Labeling in E-commerce with LLMs 17 May 2023 · 0 repositories · arXiv:2305.09858
-
M3KE: A Massive Multi-Level Multi-Subject Knowledge Evaluation Benchmark for Chinese Large Language Models 17 May 2023 · 1 repository · arXiv:2305.10263
-
Smaller Language Models are Better Black-box Machine-Generated Text Detectors 17 May 2023 · 1 repository · arXiv:2305.09859
-
Solving Cosine Similarity Underestimation between High Frequency Words by L2 Norm Discounting 17 May 2023 · 1 repository · arXiv:2305.10610
-
When Gradient Descent Meets Derivative-Free Optimization: A Match Made in Black-Box Scenario 17 May 2023 · 0 repositories · arXiv:2305.10013
-
A Preliminary Analysis on the Code Generation Capabilities of GPT-3.5 and Bard AI Models for Java Functions 16 May 2023 · 0 repositories · arXiv:2305.09402
-
CWTM: Leveraging Contextualized Word Embeddings from BERT for Neural Topic Modeling 16 May 2023 · 1 repository · arXiv:2305.09329
-
Measuring Dimensions of Self-Presentation in Twitter Bios and their Links to Misinformation Sharing 16 May 2023 · 1 repository · arXiv:2305.09548
-
Weight-Inherited Distillation for Task-Agnostic BERT Compression 16 May 2023 · 1 repository · arXiv:2305.09098
-
Coreference-aware Double-channel Attention Network for Multi-party Dialogue Reading Comprehension 15 May 2023 · 1 repository · arXiv:2305.08348
-
Keras GPT Copilot: Integrating the Power of Large Language Models in Deep Learning Model Development 15 May 2023 · 1 repository
-
Knowledge Rumination for Pre-trained Language Models 15 May 2023 · 1 repository · arXiv:2305.08732
-
Private Training Set Inspection in MLaaS 15 May 2023 · 0 repositories · arXiv:2305.09058
-
RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs 15 May 2023 · 1 repository · arXiv:2305.08844Syntology official (archive's flag): 1 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 6 pointer-only (licence)
-
Schema-adaptable Knowledge Graph Construction 15 May 2023 · 1 repository · arXiv:2305.08703
-
Similarity-weighted Construction of Contextualized Commonsense Knowledge Graphs for Knowledge-intense Argumentation Tasks 15 May 2023 · 1 repository · arXiv:2305.08495
-
Small Models are Valuable Plug-ins for Large Language Models 15 May 2023 · 1 repository · arXiv:2305.08848Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Text Classification via Large Language Models 15 May 2023 · 1 repository · arXiv:2305.08377Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Text2Gender: A Deep Learning Architecture for Analysis of Blogger's Age and Gender 15 May 2023 · 0 repositories · arXiv:2305.08633
-
Assessing the potential of LLM-assisted annotation for corpus-based pragmatics and discourse analysis: The case of apology 15 May 2023 · 0 repositories · arXiv:2305.08339
-
MatSci-NLP: Evaluating Scientific Language Models on Materials Science Language Tasks Using Text-to-Schema Modeling 14 May 2023 · 1 repository · arXiv:2305.08264Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction 14 May 2023 · 2 repositories · arXiv:2305.08144Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 17 harvested samples)
-
Bridging History with AI A Comparative Evaluation of GPT 3.5, GPT4, and GoogleBARD in Predictive Accuracy and Fact Checking 13 May 2023 · 0 repositories · arXiv:2305.07868
-
GPT-Sentinel: Distinguishing Human and ChatGPT Generated Content 13 May 2023 · 2 repositories · arXiv:2305.07969
-
The Machine Psychology of Cooperation: Can GPT models operationalise prompts for altruism, cooperation, competitiveness and selfishness in economic games? 13 May 2023 · 2 repositories · arXiv:2305.07970Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
PESTS: Persian_English Cross Lingual Corpus for Semantic Textual Similarity 13 May 2023 · 0 repositories · arXiv:2305.07893
-
Learning to Reason over Scene Graphs: A Case Study of Finetuning GPT-2 into a Robot Language Model for Grounded Task Planning 12 May 2023 · 0 repositories · arXiv:2305.07716
-
TinyStories: How Small Can Language Models Be and Still Speak Coherent English? 12 May 2023 · 8 repositories · arXiv:2305.07759Syntology 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 2 where Syntology's instrument failed) · 8 unverified (of 18 harvested samples)
-
When Giant Language Brains Just Aren't Enough! Domain Pizzazz with Knowledge Sparkle Dust 12 May 2023 · 0 repositories · arXiv:2305.07230
-
A General-Purpose Multilingual Document Encoder 11 May 2023 · 1 repository · arXiv:2305.07016
-
Generative Pre-trained Transformer: A Comprehensive Review on Enabling Technologies, Potential Applications, Emerging Challenges, and Future Directions 11 May 2023 · 0 repositories · arXiv:2305.10435
-
Spear Phishing With Large Language Models 11 May 2023 · 0 repositories · arXiv:2305.06972
-
On Practical Robust Reinforcement Learning: Practical Uncertainty Set and Double-Agent Algorithm 11 May 2023 · 0 repositories · arXiv:2305.06657
-
Overinformative Question Answering by Humans and Machines 11 May 2023 · 0 repositories · arXiv:2305.07151
-
Recommendation as Instruction Following: A Large Language Model Empowered Recommendation Approach 11 May 2023 · 0 repositories · arXiv:2305.07001
-
Transformers for CT Reconstruction From Monoplanar and Biplanar Radiographs 11 May 2023 · 0 repositories · arXiv:2305.06965
-
A Method to Automate the Discharge Summary Hospital Course for Neurology Patients 10 May 2023 · 0 repositories · arXiv:2305.06416
-
Bits of Grass: Does GPT already know how to write like Whitman? 10 May 2023 · 0 repositories · arXiv:2305.11064
-
Davinci the Dualist: the mind-body divide in large language models and in human learners 10 May 2023 · 0 repositories · arXiv:2305.07667
-
Enriching language models with graph-based context information to better understand textual data 10 May 2023 · 1 repository · arXiv:2305.11070
-
Generating medically-accurate summaries of patient-provider dialogue: A multi-stage approach using large language models 10 May 2023 · 0 repositories · arXiv:2305.05982
-
Benchmarking large language models for biomedical natural language processing applications and recommendations 10 May 2023 · 1 repository · arXiv:2305.16326
-
Summarizing, Simplifying, and Synthesizing Medical Evidence Using GPT-3 (with Varying Success) 10 May 2023 · 1 repository · arXiv:2305.06299
-
A Review of Vision-Language Models and their Performance on the Hateful Memes Challenge 9 May 2023 · 1 repository · arXiv:2305.06159
-
Alleviating Over-smoothing for Unsupervised Sentence Representation 9 May 2023 · 1 repository · arXiv:2305.06154Syntology official: harvested, nothing ran · 0 ran · 4 unverified (of 4 harvested samples)
-
Attack Named Entity Recognition by Entity Boundary Interference 9 May 2023 · 0 repositories · arXiv:2305.05253
-
CodeIE: Large Code Generation Models are Better Few-Shot Information Extractors 9 May 2023 · 1 repository · arXiv:2305.05711Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Detection of depression on social networks using transformers and ensembles 9 May 2023 · 1 repository · arXiv:2305.05325
-
GPT in Game Theory Experiments 9 May 2023 · 0 repositories · arXiv:2305.05516
-
GPT-NAS: Evolutionary Neural Architecture Search with the Generative Pre-Trained Model 9 May 2023 · 0 repositories · arXiv:2305.05351
-
Effects of sub-word segmentation on performance of transformer language models 9 May 2023 · 0 repositories · arXiv:2305.05480
-
StrAE: Autoencoding for Pre-Trained Embeddings using Explicit Structure 9 May 2023 · 0 repositories · arXiv:2305.05588
-
Towards an Automatic Optimisation Model Generator Assisted with Generative Pre-trained Transformer 9 May 2023 · 0 repositories · arXiv:2305.05811
-
Coherent Wave Dynamics and Language Generation of a Generative Pre-trained Transformer 8 May 2023 · 0 repositories · arXiv:2305.05061
-
Do Large Language Models Show Decision Heuristics Similar to Humans? A Case Study Using GPT-3.5 8 May 2023 · 0 repositories · arXiv:2305.04400
-
Explanation-based Finetuning Makes Models More Robust to Spurious Cues 8 May 2023 · 1 repository · arXiv:2305.04990
-
GersteinLab at MEDIQA-Chat 2023: Clinical Note Summarization from Doctor-Patient Conversations through Fine-tuning and In-context Learning 8 May 2023 · 0 repositories · arXiv:2305.05001
-
NeuroComparatives: Neuro-Symbolic Distillation of Comparative Knowledge 8 May 2023 · 1 repository · arXiv:2305.04978
-
PreCog: Exploring the Relation between Memorization and Performance in Pre-trained Language Models 8 May 2023 · 0 repositories · arXiv:2305.04673
-
Revisiting Relation Extraction in the era of Large Language Models 8 May 2023 · 0 repositories · arXiv:2305.05003
-
Unlocking Practical Applications in Legal Domain: Evaluation of GPT for Zero-Shot Semantic Annotation of Legal Texts 8 May 2023 · 0 repositories · arXiv:2305.04417
-
Vulnerability Detection Using Two-Stage Deep Learning Models 8 May 2023 · 0 repositories · arXiv:2305.09673
-
Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting 7 May 2023 · 2 repositories · arXiv:2305.04388Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Professional Certification Benchmark Dataset: The First 500 Jobs For Large Language Models 7 May 2023 · 0 repositories · arXiv:2305.05377
-
Stanford MLab at SemEval-2023 Task 10: Exploring GloVe- and Transformer-Based Methods for the Explainable Detection of Online Sexism 7 May 2023 · 0 repositories · arXiv:2305.04356
-
Artificial Neuropsychology: Are Large Language Models Developing Executive Functions? 6 May 2023 · 0 repositories · arXiv:2305.04134