Methods › General › Regularization › Weight Decay › Papers, page 54
Weight Decay
Papers archive 2025-07-28
archive papers tagged: 10,713 · with a code link: 4,533 · where Syntology ran a sample: 1,291 (1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,291 of 10,713 tagged: 1,064 with a run with no instrument failure, 227 where every run was a failure of Syntology's instrument)
Page 54 of 108: papers 5,301 to 5,400 of 10,713, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
BIM-GPT: a Prompt-Based Virtual Assistant Framework for BIM Information Retrieval 18 Apr 2023 · 0 repositories · arXiv:2304.09333
-
CancerGPT: Few-shot Drug Pair Synergy Prediction using Large Pre-trained Language Models 18 Apr 2023 · 0 repositories · arXiv:2304.10946
-
LLM-based Interaction for Content Generation: A Case Study on the Perception of Employees in an IT department 18 Apr 2023 · 0 repositories · arXiv:2304.09064
-
Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling 18 Apr 2023 · 1 repository · arXiv:2304.09145Syntology official (archive's flag): 5 ran · 5 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
An Empirical Study of Multitask Learning to Improve Open Domain Dialogue Systems 17 Apr 2023 · 1 repository · arXiv:2304.08115
-
Context-Dependent Embedding Utterance Representations for Emotion Recognition in Conversations 17 Apr 2023 · 1 repository · arXiv:2304.08216
-
From Zero to Hero: Examining the Power of Symbolic Tasks in Instruction Tuning 17 Apr 2023 · 1 repository · arXiv:2304.07995Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
InstructUIE: Multi-task Instruction Tuning for Unified Information Extraction 17 Apr 2023 · 1 repository · arXiv:2304.08085
-
Multimodal Short Video Rumor Detection System Based on Contrastive Learning 17 Apr 2023 · 0 repositories · arXiv:2304.08401
-
New Product Development (NPD) through Social Media-based Analysis by Comparing Word2Vec and BERT Word Embeddings 17 Apr 2023 · 0 repositories · arXiv:2304.08369
-
Supporting Qualitative Analysis with Large Language Models: Combining Codebook with GPT-3 for Deductive Coding 17 Apr 2023 · 0 repositories · arXiv:2304.10548
-
The MiniPile Challenge for Data-Efficient Language Models 17 Apr 2023 · 1 repository · arXiv:2304.08442
-
A Virtual Simulation-Pilot Agent for Training of Air Traffic Controllers 16 Apr 2023 · 0 repositories · arXiv:2304.07842
-
ArguGPT: evaluating, understanding and identifying argumentative essays generated by GPT models 16 Apr 2023 · 2 repositories · arXiv:2304.07666
-
Enhancing Automated Program Repair through Fine-tuning and Prompt Engineering 16 Apr 2023 · 0 repositories · arXiv:2304.07840
-
Sabiá: Portuguese Large Language Models 16 Apr 2023 · 0 repositories · arXiv:2304.07880
-
SikuGPT: A Generative Pre-trained Model for Intelligent Information Processing of Ancient Texts from the Perspective of Digital Humanities 16 Apr 2023 · 1 repository · arXiv:2304.07778
-
Towards Better Instruction Following Language Models for Chinese: Investigating the Impact of Training Data and Evaluation 16 Apr 2023 · 2 repositories · arXiv:2304.07854
-
Can ChatGPT Forecast Stock Price Movements? Return Predictability and Large Language Models 15 Apr 2023 · 0 repositories · arXiv:2304.07619
-
API-Bank: A Comprehensive Benchmark for Tool-Augmented LLMs 14 Apr 2023 · 2 repositories · arXiv:2304.08244Syntology 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples)
-
ChatGPT: Applications, Opportunities, and Threats 14 Apr 2023 · 0 repositories · arXiv:2304.09103
-
MedAlpaca -- An Open-Source Collection of Medical Conversational AI Models and Training Data 14 Apr 2023 · 1 repository · arXiv:2304.08247Syntology 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 1 pointer-only (licence)
-
SimpLex: a lexical text simplification architecture 14 Apr 2023 · 1 repository · arXiv:2304.07002
-
Stochastic Code Generation 14 Apr 2023 · 0 repositories · arXiv:2304.08243
-
Automated Mapping of CVE Vulnerability Records to MITRE CWE Weaknesses 13 Apr 2023 · 0 repositories · arXiv:2304.11130
-
ChatGPT cites the most-cited articles and journals, relying solely on Google Scholar's citation counts. As a result, AI may amplify the Matthew Effect in environmental science 13 Apr 2023 · 0 repositories · arXiv:2304.06794
-
Evaluation of Social Biases in Recent Large Pre-Trained Models 13 Apr 2023 · 0 repositories · arXiv:2304.06861
-
PGTask: Introducing the Task of Profile Generation from Dialogues 13 Apr 2023 · 1 repository · arXiv:2304.06634
-
Shall We Pretrain Autoregressive Language Models with Retrieval? A Comprehensive Study 13 Apr 2023 · 1 repository · arXiv:2304.06762
-
What does CLIP know about a red circle? Visual prompt engineering for VLMs 13 Apr 2023 · 0 repositories · arXiv:2304.06712
-
Detection of Fake Generated Scientific Abstracts 12 Apr 2023 · 1 repository · arXiv:2304.06148
-
Evaluation of ChatGPT Model for Vulnerability Detection 12 Apr 2023 · 0 repositories · arXiv:2304.07232
-
Localizing Model Behavior with Path Patching 12 Apr 2023 · 1 repository · arXiv:2304.05969Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
Active RIS-aided EH-NOMA Networks: A Deep Reinforcement Learning Approach 11 Apr 2023 · 0 repositories · arXiv:2304.12184
-
Approximating Online Human Evaluation of Social Chatbots with Prompting 11 Apr 2023 · 0 repositories · arXiv:2304.05253
-
Bayesian Optimization of Catalysis With In-Context Learning 11 Apr 2023 · 2 repositories · arXiv:2304.05341
-
Distinguishing ChatGPT(-3.5, -4)-generated and human-written papers through Japanese stylometric analysis 11 Apr 2023 · 0 repositories · arXiv:2304.05534
-
Exploring the Use of Foundation Models for Named Entity Recognition and Lemmatization Tasks in Slavic Languages 11 Apr 2023 · 0 repositories · arXiv:2304.05336
-
Multi-step Jailbreaking Privacy Attacks on ChatGPT 11 Apr 2023 · 1 repository · arXiv:2304.05197Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Towards preserving word order importance through Forced Invalidation 11 Apr 2023 · 1 repository · arXiv:2304.05221
-
Training Large Language Models Efficiently with Sparsity and Dataflow 11 Apr 2023 · 0 repositories · arXiv:2304.05511
-
Automated Reading Passage Generation with OpenAI's Large Language Model 10 Apr 2023 · 0 repositories · arXiv:2304.04616
-
Incorporating Structured Sentences with Time-enhanced BERT for Fully-inductive Temporal Relation Prediction 10 Apr 2023 · 0 repositories · arXiv:2304.04717
-
Is ChatGPT a Good Sentiment Analyzer? A Preliminary Study 10 Apr 2023 · 1 repository · arXiv:2304.04339Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
On the Possibilities of AI-Generated Text Detection 10 Apr 2023 · 0 repositories · arXiv:2304.04736
-
Are Large Language Models Ready for Healthcare? A Comparative Study on Clinical Language Understanding 9 Apr 2023 · 1 repository · arXiv:2304.05368
-
Learning to Tokenize for Generative Retrieval 9 Apr 2023 · 1 repository · arXiv:2304.04171
-
Factify 2: A Multimodal Fake News and Satire News Dataset 8 Apr 2023 · 1 repository · arXiv:2304.03897
-
FlexMoE: Scaling Large-scale Sparse Pre-trained Model Training via Dynamic Device Placement 8 Apr 2023 · 0 repositories · arXiv:2304.03946
-
GPT4Rec: A Generative Framework for Personalized Recommendation and User Interests Interpretation 8 Apr 2023 · 0 repositories · arXiv:2304.03879
-
Interpretable Multi Labeled Bengali Toxic Comments Classification using Deep Learning 8 Apr 2023 · 1 repository · arXiv:2304.04087
-
Multi-class Categorization of Reasons behind Mental Disturbance in Long Texts 8 Apr 2023 · 0 repositories · arXiv:2304.04118
-
tmn at SemEval-2023 Task 9: Multilingual Tweet Intimacy Detection using XLM-T, Google Translate, and Ensemble Learning 8 Apr 2023 · 1 repository · arXiv:2304.04054
-
Evaluating the Logical Reasoning Ability of ChatGPT and GPT-4 7 Apr 2023 · 1 repository · arXiv:2304.03439
-
Rethinking Evaluation Protocols of Visual Representations Learned via Self-supervised Learning 7 Apr 2023 · 0 repositories · arXiv:2304.03456
-
ChatGPT-Crawler: Find out if ChatGPT really knows what it's talking about 6 Apr 2023 · 0 repositories · arXiv:2304.03325
-
Deep Learning for Opinion Mining and Topic Classification of Course Reviews 6 Apr 2023 · 0 repositories · arXiv:2304.03394
-
GPT detectors are biased against non-native English writers 6 Apr 2023 · 2 repositories · arXiv:2304.02819Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Making AI Less "Thirsty": Uncovering and Addressing the Secret Water Footprint of AI Models 6 Apr 2023 · 1 repository · arXiv:2304.03271Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
Micron-BERT: BERT-based Facial Micro-Expression Recognition 6 Apr 2023 · 1 repository · arXiv:2304.03195Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Multi-label classification of open-ended questions with BERT 6 Apr 2023 · 0 repositories · arXiv:2304.02945
-
Towards Interpretable Mental Health Analysis with Large Language Models 6 Apr 2023 · 2 repositories · arXiv:2304.03347
-
Zero-Shot Next-Item Recommendation using Large Pretrained Language Models 6 Apr 2023 · 1 repository · arXiv:2304.03153Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Conceptual structure coheres in human cognition but not in large language models 5 Apr 2023 · 0 repositories · arXiv:2304.02754
-
Bengali Fake Review Detection using Semi-supervised Generative Adversarial Networks 5 Apr 2023 · 0 repositories · arXiv:2304.02739
-
Context-Aware Classification of Legal Document Pages 5 Apr 2023 · 0 repositories · arXiv:2304.02787
-
Document-Level Machine Translation with Large Language Models 5 Apr 2023 · 1 repository · arXiv:2304.02210
-
Large Language Models as Master Key: Unlocking the Secrets of Materials Science with GPT 5 Apr 2023 · 0 repositories · arXiv:2304.02213
-
Blockwise Compression of Transformer-based Models without Retraining 4 Apr 2023 · 0 repositories · arXiv:2304.01483
-
Geotechnical Parrot Tales (GPT): Harnessing Large Language Models in geotechnical engineering 4 Apr 2023 · 0 repositories · arXiv:2304.02138
-
GPT-4 to GPT-3.5: 'Hold My Scalpel' -- A Look at the Competency of OpenAI's GPT on the Plastic Surgery In-Service Training Exam 4 Apr 2023 · 0 repositories · arXiv:2304.01503
-
Improved Visual Fine-tuning with Natural Language Supervision 4 Apr 2023 · 1 repository · arXiv:2304.01489Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Is ChatGPT a Highly Fluent Grammatical Error Correction System? A Comprehensive Evaluation 4 Apr 2023 · 0 repositories · arXiv:2304.01746
-
LLM-Adapters: An Adapter Family for Parameter-Efficient Fine-Tuning of Large Language Models 4 Apr 2023 · 2 repositories · arXiv:2304.01933Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
REFINER: Reasoning Feedback on Intermediate Representations 4 Apr 2023 · 1 repository · arXiv:2304.01904Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
San-BERT: Extractive Summarization for Sanskrit Documents using BERT and it's variants 4 Apr 2023 · 0 repositories · arXiv:2304.01894
-
Summary of ChatGPT-Related Research and Perspective Towards the Future of Large Language Models 4 Apr 2023 · 0 repositories · arXiv:2304.01852
-
Detection of Homophobia & Transphobia in Dravidian Languages: Exploring Deep Learning Methods 3 Apr 2023 · 0 repositories · arXiv:2304.01241
-
GreekBART: The First Pretrained Greek Sequence-to-Sequence Model 3 Apr 2023 · 2 repositories · arXiv:2304.00869
-
Hate Speech Targets Detection in Parler using BERT 3 Apr 2023 · 1 repository · arXiv:2304.01179
-
MiniRBT: A Two-stage Distilled Small Chinese Pre-trained Model 3 Apr 2023 · 1 repository · arXiv:2304.00717
-
Safety Analysis in the Era of Large Language Models: A Case Study of STPA using ChatGPT 3 Apr 2023 · 2 repositories · arXiv:2304.01246
-
Does Human Collaboration Enhance the Accuracy of Identifying LLM-Generated Deepfake Texts? 3 Apr 2023 · 2 repositories · arXiv:2304.01002Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Classifying COVID-19 Related Tweets for Fake News Detection and Sentiment Analysis with BERT-based Models 2 Apr 2023 · 0 repositories · arXiv:2304.00636
-
LLMMaps -- A Visual Metaphor for Stratified Evaluation of Large Language Models 2 Apr 2023 · 1 repository · arXiv:2304.00457
-
The Other Side of Compression: Measuring Bias in Pruned Transformers 1 Apr 2023 · 1 repository
-
BERTino: an Italian DistilBERT model 31 Mar 2023 · 1 repository · arXiv:2303.18121
-
Evaluating GPT-4 and ChatGPT on Japanese Medical Licensing Examinations 31 Mar 2023 · 1 repository · arXiv:2303.18027
-
GPT-4 can pass the Korean National Licensing Examination for Korean Medicine Doctors 31 Mar 2023 · 0 repositories · arXiv:2303.17807
-
Extracting Thyroid Nodules Characteristics from Ultrasound Reports Using Transformer-based Natural Language Processing Methods 31 Mar 2023 · 0 repositories · arXiv:2304.00115
-
JobHam-place with smart recommend job options and candidate filtering options 31 Mar 2023 · 0 repositories · arXiv:2303.17930
-
Quick Dense Retrievers Consume KALE: Post Training Kullback Leibler Alignment of Embeddings for Asymmetrical dual encoders 31 Mar 2023 · 0 repositories · arXiv:2304.01016
-
Aligning a medium-size GPT model in English to a small closed domain in Spanish 30 Mar 2023 · 0 repositories · arXiv:2303.17649
-
Evaluation of GPT and BERT-based models on identifying protein-protein interactions in biomedical text 30 Mar 2023 · 0 repositories · arXiv:2303.17728
-
Fine-Tuning BERT with Character-Level Noise for Zero-Shot Transfer to Dialects and Closely-Related Languages 30 Mar 2023 · 0 repositories · arXiv:2303.17683
-
Humans in Humans Out: On GPT Converging Toward Common Sense in both Success and Failure 30 Mar 2023 · 0 repositories · arXiv:2303.17276
-
oBERTa: Improving Sparse Transfer Learning via improved initialization, distillation, and pruning regimes 30 Mar 2023 · 0 repositories · arXiv:2303.17612
-
Synthesis of Mathematical programs from Natural Language Specifications 30 Mar 2023 · 0 repositories · arXiv:2304.03287
-
Advances in apparent conceptual physics reasoning in GPT-4 29 Mar 2023 · 0 repositories · arXiv:2303.17012
-
AnnoLLM: Making Large Language Models to Be Better Crowdsourced Annotators 29 Mar 2023 · 2 repositories · arXiv:2303.16854Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)