Methods › General › Regularization › Weight Decay › Papers, page 48
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 48 of 108: papers 4,701 to 4,800 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Convolutional Neural Networks for Sentiment Analysis on Weibo Data: A Natural Language Processing Approach 13 Jul 2023 · 0 repositories · arXiv:2307.06540
-
Negated Complementary Commonsense using Large Language Models 13 Jul 2023 · 1 repository · arXiv:2307.06794
-
SecureFalcon: Are We There Yet in Automated Software Vulnerability Detection with LLMs? 13 Jul 2023 · 0 repositories · arXiv:2307.06616
-
Tackling Fake News in Bengali: Unraveling the Impact of Summarization vs. Augmentation on Pre-trained Language Models 13 Jul 2023 · 1 repository · arXiv:2307.06979
-
Retrieval Augmented Generation using Engineering Design Knowledge 13 Jul 2023 · 2 repositories · arXiv:2307.06985
-
Ashaar: Automatic Analysis and Generation of Arabic Poetry Using Deep Learning Approaches 12 Jul 2023 · 1 repository · arXiv:2307.06218
-
Detecting the Presence of COVID-19 Vaccination Hesitancy from South African Twitter Data Using Machine Learning 12 Jul 2023 · 0 repositories · arXiv:2307.15072
-
Distilling Large Language Models for Biomedical Knowledge Extraction: A Case Study on Adverse Drug Events 12 Jul 2023 · 0 repositories · arXiv:2307.06439
-
No Train No Gain: Revisiting Efficient Training Algorithms For Transformer-based Language Models 12 Jul 2023 · 1 repository · arXiv:2307.06440Syntology official (archive's flag): 9 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 3 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Online Laplace Model Selection Revisited 12 Jul 2023 · 0 repositories · arXiv:2307.06093
-
Prompt Generate Train (PGT): Few-shot Domain Adaption of Retrieval Augmented Generation Models for Open Book Question-Answering 12 Jul 2023 · 0 repositories · arXiv:2307.05915
-
Argumentative Segmentation Enhancement for Legal Summarization 11 Jul 2023 · 0 repositories · arXiv:2307.05081
-
DNAGPT: A Generalized Pre-trained Tool for Versatile DNA Sequence Analysis Tasks 11 Jul 2023 · 0 repositories · arXiv:2307.05628
-
Large Language Models 11 Jul 2023 · 0 repositories · arXiv:2307.05782
-
Named entity recognition using GPT for identifying comparable companies 11 Jul 2023 · 0 repositories · arXiv:2307.07420
-
Safe Reinforcement Learning for Strategic Bidding of Virtual Power Plants in Day-Ahead Markets 11 Jul 2023 · 0 repositories · arXiv:2307.05812
-
Unleashing the Emergent Cognitive Synergy in Large Language Models: A Task-Solving Agent through Multi-Persona Self-Collaboration 11 Jul 2023 · 2 repositories · arXiv:2307.05300
-
Vacaspati: A Diverse Corpus of Bangla Literature 11 Jul 2023 · 0 repositories · arXiv:2307.05083
-
AmadeusGPT: a natural language interface for interactive animal behavioral analysis 10 Jul 2023 · 1 repository · arXiv:2307.04858
-
ChatGPT for Digital Forensic Investigation: The Good, The Bad, and The Unknown 10 Jul 2023 · 1 repository · arXiv:2307.10195
-
SimpleMTOD: A Simple Language Model for Multimodal Task-Oriented Dialogue with Symbolic Scene Representation 10 Jul 2023 · 0 repositories · arXiv:2307.04907
-
Assessing the efficacy of large language models in generating accurate teacher responses 9 Jul 2023 · 0 repositories · arXiv:2307.04274
-
A Stitch in Time Saves Nine: Detecting and Mitigating Hallucinations of LLMs by Validating Low-Confidence Generation 8 Jul 2023 · 0 repositories · arXiv:2307.03987
-
Is ChatGPT a Good Personality Recognizer? A Preliminary Study 8 Jul 2023 · 0 repositories · arXiv:2307.03952
-
DWReCO at CheckThat! 2023: Enhancing Subjectivity Detection through Style-based Data Sampling 7 Jul 2023 · 1 repository · arXiv:2307.03550
-
Goal-Conditioned Predictive Coding for Offline Reinforcement Learning 7 Jul 2023 · 0 repositories · arXiv:2307.03406
-
How does AI chat change search behaviors? 7 Jul 2023 · 0 repositories · arXiv:2307.03826
-
RADAR: Robust AI-Text Detection via Adversarial Learning 7 Jul 2023 · 0 repositories · arXiv:2307.03838
-
TRAQ: Trustworthy Retrieval Augmented Question Answering via Conformal Prediction 7 Jul 2023 · 1 repository · arXiv:2307.04642
-
A Novel Site-Agnostic Multimodal Deep Learning Model to Identify Pro-Eating Disorder Content on Social Media 6 Jul 2023 · 0 repositories · arXiv:2307.06775
-
Can ChatGPT's Responses Boost Traditional Natural Language Processing? 6 Jul 2023 · 1 repository · arXiv:2307.04648
-
Improving Retrieval-Augmented Large Language Models via Data Importance Learning 6 Jul 2023 · 1 repository · arXiv:2307.03027Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
Large Language Models Empowered Autonomous Edge AI for Connected Intelligence 6 Jul 2023 · 0 repositories · arXiv:2307.02779
-
Text Alignment Is An Efficient Unified Model for Massive NLP Tasks 6 Jul 2023 · 1 repository · arXiv:2307.02729
-
UIT-Saviors at MEDVQA-GI 2023: Improving Multimodal Learning with Image Enhancement for Gastrointestinal Visual Question Answering 6 Jul 2023 · 0 repositories · arXiv:2307.02783
-
CAME: Confidence-guided Adaptive Memory Efficient Optimization 5 Jul 2023 · 2 repositories · arXiv:2307.02047Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples) · 1 pointer-only (licence)
-
Emoji Prediction in Tweets using BERT 5 Jul 2023 · 1 repository · arXiv:2307.02054
-
Evaluating the Effectiveness of Large Language Models in Representing Textual Descriptions of Geometry and Spatial Relations 5 Jul 2023 · 0 repositories · arXiv:2307.03678
-
External Reasoning: Towards Multi-Large-Language-Models Interchangeable Assistance with Human Feedback 5 Jul 2023 · 1 repository · arXiv:2307.12057
-
Hoodwinked: Deception and Cooperation in a Text-Based Game for Language Models 5 Jul 2023 · 1 repository · arXiv:2308.01404
-
Multilingual Controllable Transformer-Based Lexical Simplification 5 Jul 2023 · 1 repository · arXiv:2307.02120
-
Named Entity Inclusion in Abstractive Text Summarization 5 Jul 2023 · 0 repositories · arXiv:2307.02570
-
Open-Source LLMs for Text Annotation: A Practical Guide for Model Setting and Fine-Tuning 5 Jul 2023 · 0 repositories · arXiv:2307.02179
-
The FormAI Dataset: Generative AI in Software Security Through the Lens of Formal Verification 5 Jul 2023 · 0 repositories · arXiv:2307.02192
-
Embodied Task Planning with Large Language Models 4 Jul 2023 · 1 repository · arXiv:2307.01848Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
KDSTM: Neural Semi-supervised Topic Modeling with Knowledge Distillation 4 Jul 2023 · 0 repositories · arXiv:2307.01878Syntology 6 ran (of which 1 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Nonparametric Classification on Low Dimensional Manifolds using Overparameterized Convolutional Residual Networks 4 Jul 2023 · 0 repositories · arXiv:2307.01649
-
ALBERTI, a Multilingual Domain Specific Language Model for Poetry Analysis 3 Jul 2023 · 0 repositories · arXiv:2307.01387
-
Improving Language Plasticity via Pretraining with Active Forgetting 3 Jul 2023 · 1 repository · arXiv:2307.01163
-
Interpretability and Transparency-Driven Detection and Transformation of Textual Adversarial Examples (IT-DT) 3 Jul 2023 · 0 repositories · arXiv:2307.01225
-
Iterative Zero-Shot LLM Prompting for Knowledge Graph Construction 3 Jul 2023 · 0 repositories · arXiv:2307.01128
-
TensorGPT: Efficient Compression of Large Language Models based on Tensor-Train Decomposition 2 Jul 2023 · 0 repositories · arXiv:2307.00526
-
How far is Language Model from 100% Few-shot Named Entity Recognition in Medical Domain 1 Jul 2023 · 1 repository · arXiv:2307.00186
-
Large Language Models (GPT) for automating feedback on programming assignments 30 Jun 2023 · 0 repositories · arXiv:2307.00150
-
Meta-Reasoning: Semantics-Symbol Deconstruction for Large Language Models 30 Jun 2023 · 1 repository · arXiv:2306.17820
-
SPAE: Semantic Pyramid AutoEncoder for Multimodal Generation with Frozen LLMs 30 Jun 2023 · 0 repositories · arXiv:2306.17842
-
Stay on topic with Classifier-Free Guidance 30 Jun 2023 · 0 repositories · arXiv:2306.17806
-
Ticket-BERT: Labeling Incident Management Tickets with Language Models 30 Jun 2023 · 0 repositories · arXiv:2307.00108
-
A negation detection assessment of GPTs: analysis with the xNot360 dataset 29 Jun 2023 · 0 repositories · arXiv:2306.16638
-
Benchmarking Large Language Model Capabilities for Conditional Generation 29 Jun 2023 · 0 repositories · arXiv:2306.16793
-
Classifying Crime Types using Judgment Documents from Social Media 29 Jun 2023 · 0 repositories · arXiv:2306.17020
-
Generative AI for Programming Education: Benchmarking ChatGPT, GPT-4, and Human Tutors 29 Jun 2023 · 0 repositories · arXiv:2306.17156
-
Harnessing the Power of Hugging Face Transformers for Predicting Mental Health Disorders in Social Networks 29 Jun 2023 · 0 repositories · arXiv:2306.16891
-
Safety-Aware Task Composition for Discrete and Continuous Reinforcement Learning 29 Jun 2023 · 0 repositories · arXiv:2306.17033
-
Sparse Model Soups: A Recipe for Improved Pruning via Model Averaging 29 Jun 2023 · 1 repository · arXiv:2306.16788
-
Spectral Batch Normalization: Normalization in the Frequency Domain 29 Jun 2023 · 0 repositories · arXiv:2306.16999
-
Action and Trajectory Planning for Urban Autonomous Driving with Hierarchical Reinforcement Learning 28 Jun 2023 · 0 repositories · arXiv:2306.15968
-
An Efficient Sparse Inference Software Accelerator for Transformer-based Language Models on CPUs 28 Jun 2023 · 1 repository · arXiv:2306.16601
-
Pareto Optimal Learning for Estimating Large Language Model Errors 28 Jun 2023 · 0 repositories · arXiv:2306.16564
-
Beyond the Hype: Assessing the Performance, Trustworthiness, and Clinical Suitability of GPT3.5 28 Jun 2023 · 0 repositories · arXiv:2306.15887
-
Inferring the Goals of Communicating Agents from Actions and Instructions 28 Jun 2023 · 0 repositories · arXiv:2306.16207
-
Is ChatGPT a Biomedical Expert? -- Exploring the Zero-Shot Performance of Current GPT Models in Biomedical Tasks 28 Jun 2023 · 1 repository · arXiv:2306.16108
-
Multi-Site Clinical Federated Learning using Recursive and Attentive Models and NVFlare 28 Jun 2023 · 0 repositories · arXiv:2306.16367
-
Taqyim: Evaluating Arabic NLP Tasks Using ChatGPT Models 28 Jun 2023 · 1 repository · arXiv:2306.16322
-
Evaluating GPT-3.5 and GPT-4 on Grammatical Error Correction for Brazilian Portuguese 27 Jun 2023 · 0 repositories · arXiv:2306.15788
-
Gender Bias in BERT -- Measuring and Analysing Biases through Sentiment Rating in a Realistic Downstream Classification Task 27 Jun 2023 · 0 repositories · arXiv:2306.15298
-
Investigating Cross-Domain Behaviors of BERT in Review Understanding 27 Jun 2023 · 0 repositories · arXiv:2306.15123
-
MAT: Mixed-Strategy Game of Adversarial Training in Fine-tuning 27 Jun 2023 · 0 repositories · arXiv:2306.15826
-
SparseOptimizer: Sparsify Language Models through Moreau-Yosida Regularization and Accelerate via Compiler Co-design 27 Jun 2023 · 0 repositories · arXiv:2306.15656
-
Unleashing the Power of User Reviews: Exploring Airline Choices at Catania Airport, Italy 27 Jun 2023 · 0 repositories · arXiv:2306.15541
-
Constraint-aware and Ranking-distilled Token Pruning for Efficient Transformer Inference 26 Jun 2023 · 1 repository · arXiv:2306.14393
-
Exploring the Robustness of Large Language Models for Solving Programming Problems 26 Jun 2023 · 0 repositories · arXiv:2306.14583
-
LongCoder: A Long-Range Pre-trained Language Model for Code Completion 26 Jun 2023 · 1 repository · arXiv:2306.14893Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Addressing Cold Start Problem for End-to-end Automatic Speech Scoring 25 Jun 2023 · 0 repositories · arXiv:2306.14310
-
Interactive Design by Integrating a Large Pre-Trained Language Model and Building Information Modeling 25 Jun 2023 · 0 repositories · arXiv:2306.14165
-
Let's Do a Thought Experiment: Using Counterfactuals to Improve Moral Reasoning 25 Jun 2023 · 0 repositories · arXiv:2306.14308
-
Revolutionizing Cyber Threat Detection with Large Language Models: A privacy-preserving BERT-based Lightweight Model for IoT/IIoT Devices 25 Jun 2023 · 0 repositories · arXiv:2306.14263
-
Switch-BERT: Learning to Model Multimodal Interactions by Switching Attention and Input 25 Jun 2023 · 0 repositories · arXiv:2306.14182
-
Comparison of Pre-trained Language Models for Turkish Address Parsing 24 Jun 2023 · 0 repositories · arXiv:2306.13947
-
IERL: Interpretable Ensemble Representation Learning -- Combining CrowdSourced Knowledge and Distributed Semantic Representations 24 Jun 2023 · 0 repositories · arXiv:2306.13865
-
Is Pre-training Truly Better Than Meta-Learning? 24 Jun 2023 · 0 repositories · arXiv:2306.13841
-
L3Cube-MahaSent-MD: A Multi-domain Marathi Sentiment Analysis Dataset and Transformer Models 24 Jun 2023 · 1 repository · arXiv:2306.13888
-
Large Language Models as Sous Chefs: Revising Recipes with GPT-3 24 Jun 2023 · 1 repository · arXiv:2306.13986
-
Large Sequence Models for Sequential Decision-Making: A Survey 24 Jun 2023 · 0 repositories · arXiv:2306.13945
-
Math Word Problem Solving by Generating Linguistic Variants of Problem Statements 24 Jun 2023 · 1 repository · arXiv:2306.13899
-
My Boli: Code-mixed Marathi-English Corpora, Pretrained Language Models and Evaluation Benchmarks 24 Jun 2023 · 1 repository · arXiv:2306.14030
-
On the Uses of Large Language Models to Interpret Ambiguous Cyberattack Descriptions 24 Jun 2023 · 0 repositories · arXiv:2306.14062
-
Partitioning-Guided K-Means: Extreme Empty Cluster Resolution for Extreme Model Compression 24 Jun 2023 · 0 repositories · arXiv:2306.14031
-
LLM-Assisted Content Analysis: Using Large Language Models to Support Deductive Coding 23 Jun 2023 · 0 repositories · arXiv:2306.14924
-
Resume Information Extraction via Post-OCR Text Processing 23 Jun 2023 · 0 repositories · arXiv:2306.13775