Methods › General › Learning Rate Schedules › Linear Warmup With Cosine Annealing › Papers, page 28
Linear Warmup With Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,797 · with a code link: 1,655 · where Syntology ran a sample: 602 (490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (602 of 3,797 tagged: 490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument)
Page 28 of 38: papers 2,701 to 2,800 of 3,797, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
From Zero to Hero: Examining the Power of Symbolic Tasks in Instruction Tuning 17 Apr 2023 · 1 repository · arXiv:2304.07995Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Supporting Qualitative Analysis with Large Language Models: Combining Codebook with GPT-3 for Deductive Coding 17 Apr 2023 · 0 repositories · arXiv:2304.10548
-
ArguGPT: evaluating, understanding and identifying argumentative essays generated by GPT models 16 Apr 2023 · 2 repositories · arXiv:2304.07666
-
Enhancing Automated Program Repair through Fine-tuning and Prompt Engineering 16 Apr 2023 · 0 repositories · arXiv:2304.07840
-
Sabiá: Portuguese Large Language Models 16 Apr 2023 · 0 repositories · arXiv:2304.07880
-
SikuGPT: A Generative Pre-trained Model for Intelligent Information Processing of Ancient Texts from the Perspective of Digital Humanities 16 Apr 2023 · 1 repository · arXiv:2304.07778
-
Towards Better Instruction Following Language Models for Chinese: Investigating the Impact of Training Data and Evaluation 16 Apr 2023 · 2 repositories · arXiv:2304.07854
-
Can ChatGPT Forecast Stock Price Movements? Return Predictability and Large Language Models 15 Apr 2023 · 0 repositories · arXiv:2304.07619
-
API-Bank: A Comprehensive Benchmark for Tool-Augmented LLMs 14 Apr 2023 · 2 repositories · arXiv:2304.08244Syntology 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples)
-
ChatGPT: Applications, Opportunities, and Threats 14 Apr 2023 · 0 repositories · arXiv:2304.09103
-
MedAlpaca -- An Open-Source Collection of Medical Conversational AI Models and Training Data 14 Apr 2023 · 1 repository · arXiv:2304.08247Syntology 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 1 pointer-only (licence)
-
Stochastic Code Generation 14 Apr 2023 · 0 repositories · arXiv:2304.08243
-
ChatGPT cites the most-cited articles and journals, relying solely on Google Scholar's citation counts. As a result, AI may amplify the Matthew Effect in environmental science 13 Apr 2023 · 0 repositories · arXiv:2304.06794
-
PGTask: Introducing the Task of Profile Generation from Dialogues 13 Apr 2023 · 1 repository · arXiv:2304.06634
-
Shall We Pretrain Autoregressive Language Models with Retrieval? A Comprehensive Study 13 Apr 2023 · 1 repository · arXiv:2304.06762
-
What does CLIP know about a red circle? Visual prompt engineering for VLMs 13 Apr 2023 · 0 repositories · arXiv:2304.06712
-
Detection of Fake Generated Scientific Abstracts 12 Apr 2023 · 1 repository · arXiv:2304.06148
-
Evaluation of ChatGPT Model for Vulnerability Detection 12 Apr 2023 · 0 repositories · arXiv:2304.07232
-
Localizing Model Behavior with Path Patching 12 Apr 2023 · 1 repository · arXiv:2304.05969Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
Approximating Online Human Evaluation of Social Chatbots with Prompting 11 Apr 2023 · 0 repositories · arXiv:2304.05253
-
Bayesian Optimization of Catalysis With In-Context Learning 11 Apr 2023 · 2 repositories · arXiv:2304.05341
-
Distinguishing ChatGPT(-3.5, -4)-generated and human-written papers through Japanese stylometric analysis 11 Apr 2023 · 0 repositories · arXiv:2304.05534
-
Multi-step Jailbreaking Privacy Attacks on ChatGPT 11 Apr 2023 · 1 repository · arXiv:2304.05197Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Training Large Language Models Efficiently with Sparsity and Dataflow 11 Apr 2023 · 0 repositories · arXiv:2304.05511
-
Automated Reading Passage Generation with OpenAI's Large Language Model 10 Apr 2023 · 0 repositories · arXiv:2304.04616
-
On the Possibilities of AI-Generated Text Detection 10 Apr 2023 · 0 repositories · arXiv:2304.04736
-
Are Large Language Models Ready for Healthcare? A Comparative Study on Clinical Language Understanding 9 Apr 2023 · 1 repository · arXiv:2304.05368
-
GPT4Rec: A Generative Framework for Personalized Recommendation and User Interests Interpretation 8 Apr 2023 · 0 repositories · arXiv:2304.03879
-
ChatGPT-Crawler: Find out if ChatGPT really knows what it's talking about 6 Apr 2023 · 0 repositories · arXiv:2304.03325
-
GPT detectors are biased against non-native English writers 6 Apr 2023 · 2 repositories · arXiv:2304.02819Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Making AI Less "Thirsty": Uncovering and Addressing the Secret Water Footprint of AI Models 6 Apr 2023 · 1 repository · arXiv:2304.03271Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
Towards Interpretable Mental Health Analysis with Large Language Models 6 Apr 2023 · 2 repositories · arXiv:2304.03347
-
Zero-Shot Next-Item Recommendation using Large Pretrained Language Models 6 Apr 2023 · 1 repository · arXiv:2304.03153Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Conceptual structure coheres in human cognition but not in large language models 5 Apr 2023 · 0 repositories · arXiv:2304.02754
-
Document-Level Machine Translation with Large Language Models 5 Apr 2023 · 1 repository · arXiv:2304.02210
-
Large Language Models as Master Key: Unlocking the Secrets of Materials Science with GPT 5 Apr 2023 · 0 repositories · arXiv:2304.02213
-
Blockwise Compression of Transformer-based Models without Retraining 4 Apr 2023 · 0 repositories · arXiv:2304.01483
-
Geotechnical Parrot Tales (GPT): Harnessing Large Language Models in geotechnical engineering 4 Apr 2023 · 0 repositories · arXiv:2304.02138
-
GPT-4 to GPT-3.5: 'Hold My Scalpel' -- A Look at the Competency of OpenAI's GPT on the Plastic Surgery In-Service Training Exam 4 Apr 2023 · 0 repositories · arXiv:2304.01503
-
Is ChatGPT a Highly Fluent Grammatical Error Correction System? A Comprehensive Evaluation 4 Apr 2023 · 0 repositories · arXiv:2304.01746
-
LLM-Adapters: An Adapter Family for Parameter-Efficient Fine-Tuning of Large Language Models 4 Apr 2023 · 2 repositories · arXiv:2304.01933Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
REFINER: Reasoning Feedback on Intermediate Representations 4 Apr 2023 · 1 repository · arXiv:2304.01904Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Summary of ChatGPT-Related Research and Perspective Towards the Future of Large Language Models 4 Apr 2023 · 0 repositories · arXiv:2304.01852
-
GreekBART: The First Pretrained Greek Sequence-to-Sequence Model 3 Apr 2023 · 2 repositories · arXiv:2304.00869
-
Does Human Collaboration Enhance the Accuracy of Identifying LLM-Generated Deepfake Texts? 3 Apr 2023 · 2 repositories · arXiv:2304.01002Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
LLMMaps -- A Visual Metaphor for Stratified Evaluation of Large Language Models 2 Apr 2023 · 1 repository · arXiv:2304.00457
-
Evaluating GPT-4 and ChatGPT on Japanese Medical Licensing Examinations 31 Mar 2023 · 1 repository · arXiv:2303.18027
-
GPT-4 can pass the Korean National Licensing Examination for Korean Medicine Doctors 31 Mar 2023 · 0 repositories · arXiv:2303.17807
-
Aligning a medium-size GPT model in English to a small closed domain in Spanish 30 Mar 2023 · 0 repositories · arXiv:2303.17649
-
Evaluation of GPT and BERT-based models on identifying protein-protein interactions in biomedical text 30 Mar 2023 · 0 repositories · arXiv:2303.17728
-
Humans in Humans Out: On GPT Converging Toward Common Sense in both Success and Failure 30 Mar 2023 · 0 repositories · arXiv:2303.17276
-
Synthesis of Mathematical programs from Natural Language Specifications 30 Mar 2023 · 0 repositories · arXiv:2304.03287
-
Advances in apparent conceptual physics reasoning in GPT-4 29 Mar 2023 · 0 repositories · arXiv:2303.17012
-
AnnoLLM: Making Large Language Models to Be Better Crowdsourced Annotators 29 Mar 2023 · 2 repositories · arXiv:2303.16854Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
AutoAD: Movie Description in Context 29 Mar 2023 · 1 repository · arXiv:2303.16899Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 8 unverified (of 13 harvested samples)
-
Evaluating GPT-3.5 and GPT-4 Models on Brazilian University Admission Exams 29 Mar 2023 · 1 repository · arXiv:2303.17003Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 6 pointer-only (licence)
-
How do decoding algorithms distribute information in dialogue responses? 29 Mar 2023 · 0 repositories · arXiv:2303.17006
-
Ten Quick Tips for Harnessing the Power of ChatGPT/GPT-4 in Computational Biology 29 Mar 2023 · 1 repository · arXiv:2303.16429
-
ViewRefer: Grasp the Multi-view Knowledge for 3D Visual Grounding with GPT and Prototype Guidance 29 Mar 2023 · 7 repositories · arXiv:2303.16894Syntology official (archive's flag): 4 ran · 8 ran (of which 2 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Improving Large Language Models for Clinical Named Entity Recognition via Prompt Engineering 29 Mar 2023 · 1 repository · arXiv:2303.16416
-
Explicit Planning Helps Language Models in Logical Reasoning 28 Mar 2023 · 2 repositories · arXiv:2303.15714Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
On Codex Prompt Engineering for OCL Generation: An Empirical Study 28 Mar 2023 · 0 repositories · arXiv:2303.16244
-
Zero-Shot Generalizable End-to-End Task-Oriented Dialog System using Context Summarization and Domain Schema 28 Mar 2023 · 1 repository · arXiv:2303.16252
-
KPEval: Towards Fine-Grained Semantic-Based Keyphrase Evaluation 27 Mar 2023 · 1 repository · arXiv:2303.15422
-
Analyzing the Performance of GPT-3.5 and GPT-4 in Grammatical Error Correction 25 Mar 2023 · 0 repositories · arXiv:2303.14342
-
Can Large Language Models assist in Hazard Analysis? 25 Mar 2023 · 0 repositories · arXiv:2303.15473
-
GPT is becoming a Turing machine: Here are some ways to program it 25 Mar 2023 · 0 repositories · arXiv:2303.14310
-
"Get ready for a party": Exploring smarter smart spaces with help from large language models 24 Mar 2023 · 1 repository · arXiv:2303.14143
-
Personalizing Task-oriented Dialog Systems via Zero-shot Generalizable Reward Function 24 Mar 2023 · 0 repositories · arXiv:2303.13797
-
SEAL: Semantic Frame Execution And Localization for Perceiving Afforded Robot Actions 24 Mar 2023 · 0 repositories · arXiv:2303.14067
-
GesGPT: Speech Gesture Synthesis With Text Parsing from ChatGPT 23 Mar 2023 · 0 repositories · arXiv:2303.13013
-
Generate labeled training data using Prompt Programming and GPT-3. An example of Big Five Personality Classification 22 Mar 2023 · 0 repositories · arXiv:2303.12279
-
A Complete Survey on Generative AI (AIGC): Is ChatGPT from GPT-4 to GPT-5 All You Need? 21 Mar 2023 · 0 repositories · arXiv:2303.11717
-
ChatGPT and a New Academic Reality: Artificial Intelligence-Written Research Papers and the Ethics of the Large Language Models in Scholarly Publishing 21 Mar 2023 · 0 repositories · arXiv:2303.13367
-
cTBLS: Augmenting Large Language Models with Conversational Tables 21 Mar 2023 · 1 repository · arXiv:2303.12024
-
Learning A Sparse Transformer Network for Effective Image Deraining 21 Mar 2023 · 1 repository · arXiv:2303.11950Syntology official (archive's flag): 8 ran · 8 ran (of which 8 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified; every one of the 8 samples that ran constructed an object rather than computing a result (of 11 harvested samples) · 11 pointer-only (licence)
-
Sparse-IFT: Sparse Iso-FLOP Transformations for Maximizing Training Efficiency 21 Mar 2023 · 2 repositories · arXiv:2303.11525Syntology official (archive's flag): 21 ran · 21 ran (of which 0 constructed an object rather than computing a result; 20 with no instrument failure: 0 honoured, 0 violated, 20 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 24 harvested samples) · 2 pointer-only (licence)
-
Capabilities of GPT-4 on Medical Challenge Problems 20 Mar 2023 · 1 repository · arXiv:2303.13375
-
Mind meets machine: Unravelling GPT-4's cognitive psychology 20 Mar 2023 · 0 repositories · arXiv:2303.11436
-
A Comprehensive Capability Analysis of GPT-3 and GPT-3.5 Series Models 18 Mar 2023 · 0 repositories · arXiv:2303.10420
-
SPDF: Sparse Pre-training and Dense Fine-tuning for Large Language Models 18 Mar 2023 · 0 repositories · arXiv:2303.10464
-
GPTs are GPTs: An Early Look at the Labor Market Impact Potential of Large Language Models 17 Mar 2023 · 0 repositories · arXiv:2303.10130
-
Block-wise Bit-Compression of Transformer-based Models 16 Mar 2023 · 0 repositories · arXiv:2303.09184
-
Can Generative Pre-trained Transformers (GPT) Pass Assessments in Higher Education Programming Courses? 16 Mar 2023 · 0 repositories · arXiv:2303.09325
-
Jump to Conclusions: Short-Cutting Transformers With Linear Transformations 16 Mar 2023 · 2 repositories · arXiv:2303.09435Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Towards the Scalable Evaluation of Cooperativeness in Language Models 16 Mar 2023 · 0 repositories · arXiv:2303.13360
-
Automated Interactive Domain-Specific Conversational Agents that Understand Human Dialogs 15 Mar 2023 · 0 repositories · arXiv:2303.08941
-
GCRE-GPT: A Generative Model for Comparative Relation Extraction 15 Mar 2023 · 0 repositories · arXiv:2303.08601
-
SelfCheckGPT: Zero-Resource Black-Box Hallucination Detection for Generative Large Language Models 15 Mar 2023 · 1 repository · arXiv:2303.08896Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
Can ChatGPT Replace Traditional KBQA Models? An In-depth Analysis of the Question Answering Performance of the GPT LLM Family 14 Mar 2023 · 2 repositories · arXiv:2303.07992
-
RE-MOVE: An Adaptive Policy Design for Robotic Navigation Tasks in Dynamic Environments via Language-Based Feedback 14 Mar 2023 · 0 repositories · arXiv:2303.07622
-
Large Language Models in the Workplace: A Case Study on Prompt Engineering for Job Type Classification 13 Mar 2023 · 0 repositories · arXiv:2303.07142
-
Transformer-based World Models Are Happy With 100k Interactions 13 Mar 2023 · 1 repository · arXiv:2303.07109Syntology official (archive's flag): 16 ran · 16 ran (of which 6 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 1 violated, 6 with no contract checked; 8 where Syntology's instrument failed) · 9 unverified (of 25 harvested samples)
-
Large Language Models Know Your Contextual Search Intent: A Prompting Framework for Conversational Search 12 Mar 2023 · 2 repositories · arXiv:2303.06573
-
Learning Combinatorial Prompts for Universal Controllable Image Captioning 11 Mar 2023 · 0 repositories · arXiv:2303.06338
-
Algorithmic Ghost in the Research Shell: Large Language Models and Academic Knowledge Creation in Management Research 10 Mar 2023 · 0 repositories · arXiv:2303.07304
-
ChatGPT may Pass the Bar Exam soon, but has a Long Way to Go for the LexGLUE benchmark 9 Mar 2023 · 1 repository · arXiv:2304.12202
-
ICL-D3IE: In-Context Learning with Diverse Demonstrations Updating for Document Information Extraction 9 Mar 2023 · 1 repository · arXiv:2303.05063
-
Large Language Models (GPT) Struggle to Answer Multiple-Choice Questions about Code 9 Mar 2023 · 0 repositories · arXiv:2303.08033
-
ChatGPT Participates in a Computer Science Exam 8 Mar 2023 · 1 repository · arXiv:2303.09461