Methods › General › Learning Rate Schedules › Linear Warmup With Cosine Annealing › Papers, page 25
Linear Warmup With Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,797 · with a code link: 1,655 · where Syntology ran a sample: 602 (490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (602 of 3,797 tagged: 490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument)
Page 25 of 38: papers 2,401 to 2,500 of 3,797, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
InRank: Incremental Low-Rank Learning 20 Jun 2023 · 1 repository · arXiv:2306.11250
-
Learning to Generate Better Than Your LLM 20 Jun 2023 · 1 repository · arXiv:2306.11816
-
Textbooks Are All You Need 20 Jun 2023 · 0 repositories · arXiv:2306.11644
-
A Preliminary Study of ChatGPT on News Recommendation: Personalization, Provider Fairness, Fake News 19 Jun 2023 · 1 repository · arXiv:2306.10702
-
BayLing: Bridging Cross-lingual Alignment and Instruction Following through Interactive Translation for Large Language Models 19 Jun 2023 · 1 repository · arXiv:2306.10968
-
Generative Sequential Recommendation with GPTRec 19 Jun 2023 · 0 repositories · arXiv:2306.11114
-
SynerGPT: In-Context Learning for Personalized Drug Synergy Prediction and Drug Design 19 Jun 2023 · 0 repositories · arXiv:2307.11694
-
Enhancing social network hate detection using back translation and GPT-3 augmentations during training and test-time 17 Jun 2023 · 1 repository
-
Is Self-Repair a Silver Bullet for Code Generation? 16 Jun 2023 · 1 repository · arXiv:2306.09896
-
GPT4 is Slightly Helpful for Peer-Review Assistance: A Pilot Study 16 Jun 2023 · 2 repositories · arXiv:2307.05492
-
ChessGPT: Bridging Policy Learning and Language Modeling 15 Jun 2023 · 1 repository · arXiv:2306.09200Syntology official (archive's flag): 14 ran · 14 ran (of which 8 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 1 violated, 9 with no contract checked; 4 where Syntology's instrument failed) · 6 unverified (of 20 harvested samples)
-
Explore, Establish, Exploit: Red Teaming Language Models from Scratch 15 Jun 2023 · 3 repositories · arXiv:2306.09442Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Exploring the MIT Mathematics and EECS Curriculum Using Large Language Models 15 Jun 2023 · 0 repositories · arXiv:2306.08997
-
The pop song generator: designing an online course to teach collaborative, creative AI 15 Jun 2023 · 0 repositories · arXiv:2306.10069
-
Thrilled by Your Progress! Large Language Models (GPT-4) No Longer Struggle to Pass Assessments in Higher Education Programming Courses 15 Jun 2023 · 0 repositories · arXiv:2306.10073
-
Assessing the Effectiveness of GPT-3 in Detecting False Political Statements: A Case Study on the LIAR Dataset 14 Jun 2023 · 1 repository · arXiv:2306.08190
-
Language models are not naysayers: An analysis of language models on negation benchmarks 14 Jun 2023 · 1 repository · arXiv:2306.08189
-
Towards AGI in Computer Vision: Lessons Learned from GPT and Large Language Models 14 Jun 2023 · 0 repositories · arXiv:2306.08641
-
Enhancing Social Network Hate Detection Using Back Translation and GPT-3 Augmentations During Training and Test-Time 13 Jun 2023 · 1 repository
-
FLamE: Few-shot Learning from Natural Language Explanations 13 Jun 2023 · 0 repositories · arXiv:2306.08042
-
Human-Like Intuitive Behavior and Reasoning Biases Emerged in Language Models -- and Disappeared in GPT-4 13 Jun 2023 · 0 repositories · arXiv:2306.07622
-
A Survey of Vision-Language Pre-training from the Lens of Multimodal Machine Translation 12 Jun 2023 · 0 repositories · arXiv:2306.07198
-
On the N-gram Approximation of Pre-trained Language Models 12 Jun 2023 · 0 repositories · arXiv:2306.06892
-
Recursion of Thought: A Divide-and-Conquer Approach to Multi-Context Reasoning with Language Models 12 Jun 2023 · 1 repository · arXiv:2306.06891
-
The BEA 2023 Shared Task on Generating AI Teacher Responses in Educational Dialogues 12 Jun 2023 · 0 repositories · arXiv:2306.06941
-
Waffling around for Performance: Visual Classification with Random Words and Broad Concepts 12 Jun 2023 · 2 repositories · arXiv:2306.07282Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Inductive reasoning in humans and large language models 11 Jun 2023 · 1 repository · arXiv:2306.06548
-
Exploring the Responses of Large Language Models to Beginner Programmers' Help Requests 9 Jun 2023 · 0 repositories · arXiv:2306.05715
-
GPT-Calls: Enhancing Call Segmentation and Tagging by Generating Synthetic Conversations via Large Language Models 9 Jun 2023 · 0 repositories · arXiv:2306.07941
-
Language Models Can Learn Exceptions to Syntactic Rules 9 Jun 2023 · 1 repository · arXiv:2306.05969
-
Prodigy: An Expeditiously Adaptive Parameter-Free Learner 9 Jun 2023 · 1 repository · arXiv:2306.06101
-
Reliability Check: An Analysis of GPT-3's Response to Sensitive Topics and Prompt Wording 9 Jun 2023 · 2 repositories · arXiv:2306.06199
-
Understanding Telecom Language Through Large Language Models 9 Jun 2023 · 0 repositories · arXiv:2306.07933
-
PandaLM: An Automatic Evaluation Benchmark for LLM Instruction Tuning Optimization 8 Jun 2023 · 2 repositories · arXiv:2306.05087Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Prefer to Classify: Improving Text Classifiers via Auxiliary Preference Learning 8 Jun 2023 · 1 repository · arXiv:2306.04925Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
The ADAIO System at the BEA-2023 Shared Task on Generating AI Teacher Responses in Educational Dialogues 8 Jun 2023 · 0 repositories · arXiv:2306.05360
-
ToolAlpaca: Generalized Tool Learning for Language Models with 3000 Simulated Cases 8 Jun 2023 · 3 repositories · arXiv:2306.05301Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples)
-
On the Detectability of ChatGPT Content: Benchmarking, Methodology, and Evaluation through the Lens of Academic Writing 7 Jun 2023 · 2 repositories · arXiv:2306.05524
-
Good Data, Large Data, or No Data? Comparing Three Approaches in Developing Research Aspect Classifiers for Biomedical Papers 7 Jun 2023 · 1 repository · arXiv:2306.04820
-
GPT Self-Supervision for a Better Data Annotator 7 Jun 2023 · 0 repositories · arXiv:2306.04349
-
Personality testing of Large Language Models: Limited temporal stability, but highlighted prosociality 7 Jun 2023 · 0 repositories · arXiv:2306.04308
-
ScienceBenchmark: A Complex Real-World Benchmark for Evaluating Natural Language to SQL Systems 7 Jun 2023 · 0 repositories · arXiv:2306.04743
-
The Two Word Test: A Semantic Benchmark for Large Language Models 7 Jun 2023 · 1 repository · arXiv:2306.04610
-
An Empirical Analysis of Parameter-Efficient Methods for Debiasing Pre-Trained Language Models 6 Jun 2023 · 1 repository · arXiv:2306.04067
-
Certified Deductive Reasoning with Language Models 6 Jun 2023 · 1 repository · arXiv:2306.04031Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Iterative Translation Refinement with Large Language Models 6 Jun 2023 · 0 repositories · arXiv:2306.03856
-
Language acquisition: do children and language models follow similar learning stages? 6 Jun 2023 · 0 repositories · arXiv:2306.03586
-
Analyzing Syntactic Generalization Capacity of Pre-trained Language Models on Japanese Honorific Conversion 5 Jun 2023 · 0 repositories · arXiv:2306.03055
-
ChatGPT as a mapping assistant: A novel method to enrich maps with generative AI and content derived from street-level photographs 5 Jun 2023 · 0 repositories · arXiv:2306.03204
-
Efficient GPT Model Pre-training using Tensor Train Matrix Representation 5 Jun 2023 · 0 repositories · arXiv:2306.02697
-
Skill over Scale: The Case for Medium, Domain-Specific Models for SE 5 Jun 2023 · 0 repositories · arXiv:2306.03268
-
Auto-GPT for Online Decision Making: Benchmarks and Additional Opinions 4 Jun 2023 · 1 repository · arXiv:2306.02224Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples) · 4 pointer-only (licence)
-
Towards Coding Social Science Datasets with Language Models 3 Jun 2023 · 0 repositories · arXiv:2306.02177
-
Can Contextual Biasing Remain Effective with Whisper and GPT-2? 2 Jun 2023 · 1 repository · arXiv:2306.01942
-
Automatic Glossary of Clinical Terminology: a Large-Scale Dictionary of Biomedical Definitions Generated from Ontological Knowledge 1 Jun 2023 · 0 repositories · arXiv:2306.00665
-
Enhancing Programming eTextbooks with ChatGPT Generated Counterfactual-Thinking-Inspired Questions 1 Jun 2023 · 0 repositories · arXiv:2306.00551
-
Multi-Dimensional Evaluation of Text Summarization with In-Context Learning 1 Jun 2023 · 1 repository · arXiv:2306.01200
-
Systematic Evaluation of GPT-3 for Zero-Shot Personality Estimation 1 Jun 2023 · 0 repositories · arXiv:2306.01183
-
TopEx: Topic-based Explanations for Model Comparison 1 Jun 2023 · 0 repositories · arXiv:2306.00976
-
Evaluating GPT's Programming Capability through CodeWars' Katas 31 May 2023 · 0 repositories · arXiv:2306.01784
-
Examining the Emergence of Deductive Reasoning in Generative Language Models 31 May 2023 · 0 repositories · arXiv:2306.01009
-
Harnessing Explanations: LLM-to-LM Interpreter for Enhanced Text-Attributed Graph Representation Learning 31 May 2023 · 3 repositories · arXiv:2305.19523Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Knowledge Base Question Answering for Space Debris Queries 31 May 2023 · 1 repository · arXiv:2305.19734
-
Does Conceptual Representation Require Embodiment? Insights From Large Language Models 30 May 2023 · 0 repositories · arXiv:2305.19103
-
Generate then Select: Open-ended Visual Question Answering Guided by World Knowledge 30 May 2023 · 0 repositories · arXiv:2305.18842
-
GPT Models in Construction Industry: Opportunities, Limitations, and a Use Case Validation 30 May 2023 · 0 repositories · arXiv:2305.18997
-
Seeing Seeds Beyond Weeds: Green Teaming Generative AI for Beneficial Uses 30 May 2023 · 0 repositories · arXiv:2306.03097
-
Check-COVID: Fact-Checking COVID-19 News Claims with Scientific Evidence 29 May 2023 · 1 repository · arXiv:2305.18265
-
Coeditor: Leveraging Contextual Changes for Multi-round Code Auto-editing 29 May 2023 · 0 repositories · arXiv:2305.18584
-
Do Large Language Models Know What They Don't Know? 29 May 2023 · 1 repository · arXiv:2305.18153Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Exploring Effectiveness of GPT-3 in Grammatical Error Correction: A Study on Performance and Controllability in Prompt-Based Methods 29 May 2023 · 0 repositories · arXiv:2305.18156
-
LM-CPPF: Paraphrasing-Guided Data Augmentation for Contrastive Prompt-Based Few-Shot Fine-Tuning 29 May 2023 · 1 repository · arXiv:2305.18169
-
Marked Personas: Using Natural Language Prompts to Measure Stereotypes in Language Models 29 May 2023 · 1 repository · arXiv:2305.18189Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
ProcessGPT: Transforming Business Process Management with Generative Artificial Intelligence 29 May 2023 · 0 repositories · arXiv:2306.01771
-
Syntax and Semantics Meet in the "Middle": Probing the Syntax-Semantics Interface of LMs Through Agentivity 29 May 2023 · 1 repository · arXiv:2305.18185
-
Test-Time Training on Nearest Neighbors for Large Language Models 29 May 2023 · 1 repository · arXiv:2305.18466Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Transformer Language Models Handle Word Frequency in Prediction Head 29 May 2023 · 0 repositories · arXiv:2305.18294
-
Bridging the Language Gap: Dynamic Learning Strategies for Improving Multilingual Performance in LLMs 28 May 2023 · 0 repositories · arXiv:2305.17740
-
Evaluating GPT-3 Generated Explanations for Hateful Content Moderation 28 May 2023 · 1 repository · arXiv:2305.17680
-
Generating EDU Extracts for Plan-Guided Summary Re-Ranking 28 May 2023 · 1 repository · arXiv:2305.17779Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Knowledge-Augmented Reasoning Distillation for Small Language Models in Knowledge-Intensive Tasks 28 May 2023 · 1 repository · arXiv:2305.18395Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
KoSBi: A Dataset for Mitigating Social Bias Risks Towards Safer Large Language Model Application 28 May 2023 · 1 repository · arXiv:2305.17701
-
Mitigating Label Biases for In-context Learning 28 May 2023 · 1 repository · arXiv:2305.19148Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
SQuARe: A Large-Scale Dataset of Sensitive Questions and Acceptable Responses Created Through Human-Machine Collaboration 28 May 2023 · 1 repository · arXiv:2305.17696
-
Transfer Learning for Power Outage Detection Task with Limited Training Data 28 May 2023 · 0 repositories · arXiv:2305.17817
-
Complementary and Integrative Health Lexicon (CIHLex) and Entity Recognition in the Literature 27 May 2023 · 0 repositories · arXiv:2305.17353
-
DNA-GPT: Divergent N-Gram Analysis for Training-Free Detection of GPT-Generated Text 27 May 2023 · 1 repository · arXiv:2305.17359Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
The Curse of Recursion: Training on Generated Data Makes Models Forget 27 May 2023 · 1 repository · arXiv:2305.17493
-
Towards Explainable Conversational Recommender Systems 27 May 2023 · 1 repository · arXiv:2305.18363
-
What can Large Language Models do in chemistry? A comprehensive benchmark on eight tasks 27 May 2023 · 1 repository · arXiv:2305.18365
-
Backpack Language Models 26 May 2023 · 1 repository · arXiv:2305.16765Syntology 3 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; every one of the 3 samples that ran constructed an object rather than computing a result (of 4 harvested samples)
-
Chain-of-Thought Hub: A Continuous Effort to Measure Large Language Models' Reasoning Performance 26 May 2023 · 1 repository · arXiv:2305.17306Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
ChatGPT: A Study on its Utility for Ubiquitous Software Engineering Tasks 26 May 2023 · 0 repositories · arXiv:2305.16837
-
Counterfactual reasoning: Testing language models' understanding of hypothetical scenarios 26 May 2023 · 1 repository · arXiv:2305.16572
-
Distinguishing Human Generated Text From ChatGPT Generated Text Using Machine Learning 26 May 2023 · 0 repositories · arXiv:2306.01761
-
Do GPTs Produce Less Literal Translations? 26 May 2023 · 1 repository · arXiv:2305.16806
-
Evaluation of Question Generation Needs More References 26 May 2023 · 0 repositories · arXiv:2305.16626
-
Impossible Distillation: from Low-Quality Model to High-Quality Dataset & Model for Summarization and Paraphrasing 26 May 2023 · 0 repositories · arXiv:2305.16635
-
Improving accuracy of GPT-3/4 results on biomedical data using a retrieval-augmented language model 26 May 2023 · 0 repositories · arXiv:2305.17116
-
Large Language Models as Tool Makers 26 May 2023 · 1 repository · arXiv:2305.17126