Methods › General › Regularization › Weight Decay › Papers, page 56
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 56 of 108: papers 5,501 to 5,600 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Prompt, Generate, then Cache: Cascade of Foundation Models makes Strong Few-shot Learners 3 Mar 2023 · 3 repositories · arXiv:2303.02151Syntology official: not harvested · 0 ran · 1 unverified (of 1 harvested sample)
-
Prophet: Prompting Large Language Models with Complementary Answer Heuristics for Knowledge-based Visual Question Answering 3 Mar 2023 · 1 repository · arXiv:2303.01903Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
TrojText: Test-time Invisible Textual Trojan Insertion 3 Mar 2023 · 1 repository · arXiv:2303.02242Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Will Affective Computing Emerge from Foundation Models and General AI? A First Evaluation on ChatGPT 3 Mar 2023 · 0 repositories · arXiv:2303.03186
-
Adopting the Multi-answer Questioning Task with an Auxiliary Metric for Extreme Multi-label Text Classification Utilizing the Label Hierarchy 2 Mar 2023 · 0 repositories · arXiv:2303.01064
-
Can BERT Refrain from Forgetting on Sequential Tasks? A Probing Study 2 Mar 2023 · 1 repository · arXiv:2303.01081
-
Evaluating Parameter-Efficient Transfer Learning Approaches on SURE Benchmark for Speech Understanding 2 Mar 2023 · 1 repository · arXiv:2303.03267
-
INO at Factify 2: Structure Coherence based Multi-Modal Fact Verification 2 Mar 2023 · 1 repository · arXiv:2303.01510
-
Sparse MoE as the New Dropout: Scaling Dense and Self-Slimmable Transformers 2 Mar 2023 · 1 repository · arXiv:2303.01610Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
WiCE: Real-World Entailment for Claims in Wikipedia 2 Mar 2023 · 2 repositories · arXiv:2303.01432Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
A Framework for Neurosymbolic Robot Action Planning using Large Language Models 1 Mar 2023 · 1 repository · arXiv:2303.00438Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Competence-Based Analysis of Language Models 1 Mar 2023 · 0 repositories · arXiv:2303.00333
-
Domain-adapted large language models for classifying nuclear medicine reports 1 Mar 2023 · 0 repositories · arXiv:2303.01258
-
How Robust is GPT-3.5 to Predecessors? A Comprehensive Study on Language Understanding Tasks 1 Mar 2023 · 0 repositories · arXiv:2303.00293
-
ToxVis: Enabling Interpretability of Implicit vs. Explicit Toxicity Detection Models with Interactive Visualization 1 Mar 2023 · 0 repositories · arXiv:2303.09402
-
Automatically Classifying Emotions based on Text: A Comparative Exploration of Different Datasets 28 Feb 2023 · 0 repositories · arXiv:2302.14727
-
Zero-Shot Cross-Lingual Summarization via Large Language Models 28 Feb 2023 · 0 repositories · arXiv:2302.14229
-
Information-Restricted Neural Language Models Reveal Different Brain Regions' Sensitivity to Semantics, Syntax and Context 28 Feb 2023 · 1 repository · arXiv:2302.14389Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Large Language Models Are State-of-the-Art Evaluators of Translation Quality 28 Feb 2023 · 4 repositories · arXiv:2302.14520
-
Sampled Transformer for Point Sets 28 Feb 2023 · 0 repositories · arXiv:2302.14346
-
Text classification dataset and analysis for Uzbek language 28 Feb 2023 · 1 repository · arXiv:2302.14494
-
Weighted Sampling for Masked Language Modeling 28 Feb 2023 · 0 repositories · arXiv:2302.14225
-
Combating Uncertainties in Wind and Distributed PV Energy Sources Using Integrated Reinforcement Learning and Time-Series Forecasting 27 Feb 2023 · 0 repositories · arXiv:2302.14094
-
Elementwise Language Representation 27 Feb 2023 · 0 repositories · arXiv:2302.13475
-
Inseq: An Interpretability Toolkit for Sequence Generation Models 27 Feb 2023 · 2 repositories · arXiv:2302.13942
-
LLaMA: Open and Efficient Foundation Language Models 27 Feb 2023 · 57 repositories · arXiv:2302.13971Syntology official: no sample here; runs from other or unrecorded repositories · 37 ran (of which 9 constructed an object rather than computing a result; 25 with no instrument failure: 3 honoured, 0 violated, 22 with no contract checked; 12 where Syntology's instrument failed) · 21 unverified (of 58 harvested samples) · 4 pointer-only (licence)
-
Reward Design with Language Models 27 Feb 2023 · 1 repository · arXiv:2303.00001
-
Systematic Rectification of Language Models via Dead-end Analysis 27 Feb 2023 · 1 repository · arXiv:2302.14003Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
Using Auxiliary Tasks In Multimodal Fusion Of Wav2vec 2.0 And BERT For Multimodal Emotion Recognition 27 Feb 2023 · 0 repositories · arXiv:2302.13661
-
Comparing Sentence-Level Suggestions to Message-Level Suggestions in AI-Mediated Communication 26 Feb 2023 · 0 repositories · arXiv:2302.13382
-
Efficient Ensemble for Multimodal Punctuation Restoration using Time-Delay Neural Network 26 Feb 2023 · 1 repository · arXiv:2302.13376
-
Fast Attention Requires Bounded Entries 26 Feb 2023 · 0 repositories · arXiv:2302.13214
-
Human-in-the-Loop Schema Induction 25 Feb 2023 · 0 repositories · arXiv:2302.13048
-
HULAT at SemEval-2023 Task 10: Data augmentation for pre-trained transformers applied to the detection of sexism in social media 24 Feb 2023 · 1 repository · arXiv:2302.12840
-
MUX-PLMs: Data Multiplexing for High-throughput Language Models 24 Feb 2023 · 1 repository · arXiv:2302.12441
-
Spanish Built Factual Freectianary (Spanish-BFF): the first AI-generated free dictionary 24 Feb 2023 · 0 repositories · arXiv:2302.12746
-
Window transformer for dialogue document: a joint framework for causal emotion entailment 24 Feb 2023 · 0 repositories
-
Revisiting the Gumbel-Softmax in MADDPG 23 Feb 2023 · 1 repository · arXiv:2302.11793
-
Teacher Intervention: Improving Convergence of Quantization Aware Training for Ultra-Low Precision Transformers 23 Feb 2023 · 1 repository · arXiv:2302.11812Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Testing AI on language comprehension tasks reveals insensitivity to underlying meaning 23 Feb 2023 · 0 repositories · arXiv:2302.12313
-
What makes a language easy to deep-learn? Deep neural networks and humans similarly benefit from compositional structure 23 Feb 2023 · 1 repository · arXiv:2302.12239Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Solution for the EPO CodeFest on Green Plastics: Hierarchical multi-label classification of patents relating to green plastics using deep learning 22 Feb 2023 · 0 repositories · arXiv:2302.13784
-
kNN-Adapter: Efficient Domain Adaptation for Black-Box Language Models 21 Feb 2023 · 0 repositories · arXiv:2302.10879
-
UAV Path Planning Employing MPC- Reinforcement Learning Method Considering Collision Avoidance 21 Feb 2023 · 0 repositories · arXiv:2302.10669
-
Boosting classification reliability of NLP transformer models in the long run 20 Feb 2023 · 0 repositories · arXiv:2302.10016
-
Large-scale Multi-Modal Pre-trained Models: A Comprehensive Survey 20 Feb 2023 · 1 repository · arXiv:2302.10035
-
ChatIE: Zero-Shot Information Extraction via Chatting with ChatGPT 20 Feb 2023 · 1 repository · arXiv:2302.10205
-
Can ChatGPT Understand Too? A Comparative Study on ChatGPT and Fine-tuned BERT 19 Feb 2023 · 1 repository · arXiv:2302.10198
-
Evaluating the Effectiveness of Pre-trained Language Models in Predicting the Helpfulness of Online Product Reviews 19 Feb 2023 · 1 repository · arXiv:2302.10199
-
Text Classification in the Wild: a Large-scale Long-tailed Name Normalization Dataset 19 Feb 2023 · 1 repository · arXiv:2302.09509
-
A Comprehensive Survey on Pretrained Foundation Models: A History from BERT to ChatGPT 18 Feb 2023 · 0 repositories · arXiv:2302.09419
-
How Good Are GPT Models at Machine Translation? A Comprehensive Evaluation 18 Feb 2023 · 1 repository · arXiv:2302.09210Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Bounding the Capabilities of Large Language Models in Open Text Generation with Prompt Constraints 17 Feb 2023 · 1 repository · arXiv:2302.09185
-
Conveying the Predicted Future to Users: A Case Study of Story Plot Prediction 17 Feb 2023 · 1 repository · arXiv:2302.09122
-
GPT4MIA: Utilizing Generative Pre-trained Transformer (GPT-3) as A Plug-and-Play Transductive Model for Medical Image Analysis 17 Feb 2023 · 0 repositories · arXiv:2302.08722
-
Hate Speech and Offensive Language Detection using an Emotion-aware Shared Encoder 17 Feb 2023 · 0 repositories · arXiv:2302.08777
-
PAC Prediction Sets for Large Language Models of Code 17 Feb 2023 · 1 repository · arXiv:2302.08703
-
Prompting Large Language Models With the Socratic Method 17 Feb 2023 · 0 repositories · arXiv:2303.08769
-
ViTA: A Vision Transformer Inference Accelerator for Edge Applications 17 Feb 2023 · 0 repositories · arXiv:2302.09108
-
Foundation Models for Natural Language Processing -- Pre-trained Language Models Integrating Media 16 Feb 2023 · 0 repositories · arXiv:2302.08575
-
For Generated Text, Is NLI-Neutral Text the Best Text? 16 Feb 2023 · 1 repository · arXiv:2302.08577
-
Marich: A Query-efficient Distributionally Equivalent Model Extraction Attack using Public Data 16 Feb 2023 · 1 repository · arXiv:2302.08466Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples)
-
Retrieval-augmented Image Captioning 16 Feb 2023 · 1 repository · arXiv:2302.08268
-
Syntactic Structure Processing in the Brain while Listening 16 Feb 2023 · 0 repositories · arXiv:2302.08589
-
Commonsense Reasoning for Conversational AI: A Survey of the State of the Art 15 Feb 2023 · 0 repositories · arXiv:2302.07926
-
Learning Performance-Improving Code Edits 15 Feb 2023 · 2 repositories · arXiv:2302.07867Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 5 where Syntology's instrument failed) · 11 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Smart 6G Sky for Green Mobile IOT Networks 15 Feb 2023 · 1 repository · arXiv:2302.09022
-
Towards Optimal Compression: Joint Pruning and Quantization 15 Feb 2023 · 0 repositories · arXiv:2302.07612
-
Tree-Based Representation and Generation of Natural and Mathematical Language 15 Feb 2023 · 1 repository · arXiv:2302.07974
-
A Modern Look at the Relationship between Sharpness and Generalization 14 Feb 2023 · 1 repository · arXiv:2302.07011
-
A Psycholinguistic Analysis of BERT's Representations of Compounds 14 Feb 2023 · 1 repository · arXiv:2302.07232
-
Exploring Category Structure with Contextual Language Models and Lexical Semantic Networks 14 Feb 2023 · 0 repositories · arXiv:2302.06942
-
Few-shot learning approaches for classifying low resource domain specific software requirements 14 Feb 2023 · 0 repositories · arXiv:2302.06951
-
Reveal the Unknown: Out-of-Knowledge-Base Mention Discovery with Entity Linking 14 Feb 2023 · 3 repositories · arXiv:2302.07189
-
ScatterShot: Interactive In-context Example Curation for Text Transformation 14 Feb 2023 · 1 repository · arXiv:2302.07346
-
Diminished Diversity-of-Thought in a Standard Large Language Model 13 Feb 2023 · 0 repositories · arXiv:2302.07267
-
Can GPT-3 Perform Statutory Reasoning? 13 Feb 2023 · 1 repository · arXiv:2302.06100
-
Linguistic ambiguity analysis in ChatGPT 13 Feb 2023 · 0 repositories · arXiv:2302.06426
-
STREET: A Multi-Task Structured Reasoning and Explanation Benchmark 13 Feb 2023 · 0 repositories · arXiv:2302.06729
-
Academic Writing with GPT-3.5: Reflections on Practices, Efficacy and Transparency 12 Feb 2023 · 0 repositories · arXiv:2304.11079
-
Semantic Importance-Aware Communications Using Pre-trained Language Models 12 Feb 2023 · 0 repositories · arXiv:2302.07142
-
Transformer models: an introduction and catalog 12 Feb 2023 · 0 repositories · arXiv:2302.07730
-
A Brief Report on LawGPT 1.0: A Virtual Legal Assistant Based on GPT-3 11 Feb 2023 · 0 repositories · arXiv:2302.05729
-
DocILE Benchmark for Document Information Localization and Extraction 11 Feb 2023 · 1 repository · arXiv:2302.05658
-
Informing clinical assessment by contextualizing post-hoc explanations of risk prediction models in type-2 diabetes 11 Feb 2023 · 0 repositories · arXiv:2302.05752
-
Alloprof: a new French question-answer education dataset and its use in an information retrieval case study 10 Feb 2023 · 1 repository · arXiv:2302.07738
-
BEST: BERT Pre-Training for Sign Language Recognition with Coupling Tokenization 10 Feb 2023 · 0 repositories · arXiv:2302.05075
-
Combat AI With AI: Counteract Machine-Generated Fake Restaurant Reviews on Social Media 10 Feb 2023 · 1 repository · arXiv:2302.07731
-
FairPy: A Toolkit for Evaluation of Prediction Biases and their Mitigation in Large Language Models 10 Feb 2023 · 1 repository · arXiv:2302.05508
-
GTR-CTRL: Instrument and Genre Conditioning for Guitar-Focused Music Generation with Transformers 10 Feb 2023 · 0 repositories · arXiv:2302.05393
-
RIS-Assisted Jamming Rejection and Path Planning for UAV-Borne IoT Platform: A New Deep Reinforcement Learning Framework 10 Feb 2023 · 0 repositories · arXiv:2302.04994
-
The Wisdom of Hindsight Makes Language Models Better Instruction Followers 10 Feb 2023 · 1 repository · arXiv:2302.05206
-
Translating Natural Language to Planning Goals with Large-Language Models 10 Feb 2023 · 1 repository · arXiv:2302.05128
-
Better by you, better than me, chatgpt3 as writing assistance in students essays 9 Feb 2023 · 0 repositories · arXiv:2302.04536
-
Generating a Structured Summary of Numerous Academic Papers: Dataset and Method 9 Feb 2023 · 1 repository · arXiv:2302.04580
-
CRL+: A Novel Semi-Supervised Deep Active Contrastive Representation Learning-Based Text Classification Model for Insurance Data 8 Feb 2023 · 0 repositories · arXiv:2302.04343
-
Efficient Joint Learning for Clinical Named Entity Recognition and Relation Extraction Using Fourier Networks: A Use Case in Adverse Drug Events 8 Feb 2023 · 1 repository · arXiv:2302.04185
-
Prompting for Multimodal Hateful Meme Classification 8 Feb 2023 · 0 repositories · arXiv:2302.04156
-
LUT-NN: Empower Efficient Neural Network Inference with Centroid Learning and Table Lookup 7 Feb 2023 · 0 repositories · arXiv:2302.03213
-
Reliable Natural Language Understanding with Large Language Models and Answer Set Programming 7 Feb 2023 · 0 repositories · arXiv:2302.03780