Methods › General › Learning Rate Schedules › Cosine Annealing › Papers, page 28
Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,965 · with a code link: 1,734 · where Syntology ran a sample: 627 (513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (627 of 3,965 tagged: 513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument)
Page 28 of 40: papers 2,701 to 2,800 of 3,965, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
FreeLM: Fine-Tuning-Free Language Model 2 May 2023 · 0 repositories · arXiv:2305.01616
-
How to Unleash the Power of Large Language Models for Few-shot Relation Extraction? 2 May 2023 · 2 repositories · arXiv:2305.01555
-
A Paradigm Shift: The Future of Machine Translation Lies with Large Language Models 2 May 2023 · 0 repositories · arXiv:2305.01181
-
Vision Meets Definitions: Unsupervised Visual Word Sense Disambiguation Incorporating Gloss Information 2 May 2023 · 1 repository · arXiv:2305.01788
-
Automated Paper Screening for Clinical Reviews Using Large Language Models 1 May 2023 · 0 repositories · arXiv:2305.00844
-
Beyond Classification: Financial Reasoning in State-of-the-Art Language Models 30 Apr 2023 · 1 repository · arXiv:2305.01505
-
Using Large Language Models to Generate JUnit Tests: An Empirical Study 30 Apr 2023 · 1 repository · arXiv:2305.00418
-
How does GPT-2 compute greater-than?: Interpreting mathematical abilities in a pre-trained language model 30 Apr 2023 · 3 repositories · arXiv:2305.00586
-
Causal Reasoning and Large Language Models: Opening a New Frontier for Causality 28 Apr 2023 · 1 repository · arXiv:2305.00050Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
FlowTransformer: A Transformer Framework for Flow-based Network Intrusion Detection Systems 28 Apr 2023 · 1 repository · arXiv:2304.14746
-
Towards Automated Circuit Discovery for Mechanistic Interpretability 28 Apr 2023 · 4 repositories · arXiv:2304.14997Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples)
-
Framing the News:From Human Perception to Large Language Model Inferences 27 Apr 2023 · 0 repositories · arXiv:2304.14456
-
ICE-Score: Instructing Large Language Models to Evaluate Code 27 Apr 2023 · 2 repositories · arXiv:2304.14317Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Origin Tracing and Detecting of LLMs 27 Apr 2023 · 0 repositories · arXiv:2304.14072
-
SweCTRL-Mini: a data-transparent Transformer-based large language model for controllable text generation in Swedish 27 Apr 2023 · 1 repository · arXiv:2304.13994
-
Prompting GPT-3.5 for Text-to-SQL with De-semanticization and Skeleton Retrieval 26 Apr 2023 · 0 repositories · arXiv:2304.13301
-
Evaluation of GPT-3.5 and GPT-4 for supporting real-world information needs in healthcare delivery 26 Apr 2023 · 0 repositories · arXiv:2304.13714
-
Exploiting CNNs for Semantic Segmentation with Pascal VOC 26 Apr 2023 · 0 repositories · arXiv:2304.13216
-
Exploring the Curious Case of Code Prompts 26 Apr 2023 · 1 repository · arXiv:2304.13250Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
Extracting Structured Seed-Mediated Gold Nanorod Growth Procedures from Literature with GPT-3 26 Apr 2023 · 0 repositories · arXiv:2304.13846
-
Towards Multi-Modal DBMSs for Seamless Querying of Texts and Tables 26 Apr 2023 · 0 repositories · arXiv:2304.13559
-
Measuring Massive Multitask Chinese Understanding 25 Apr 2023 · 2 repositories · arXiv:2304.12986Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Semantic Compression With Large Language Models 25 Apr 2023 · 0 repositories · arXiv:2304.12512
-
The Potential of Visual ChatGPT For Remote Sensing 25 Apr 2023 · 0 repositories · arXiv:2304.13009
-
Generation-driven Contrastive Self-training for Zero-shot Text Classification with Instruction-following LLM 24 Apr 2023 · 1 repository · arXiv:2304.11872
-
Boosting Theory-of-Mind Performance in Large Language Models via Prompting 22 Apr 2023 · 1 repository · arXiv:2304.11490
-
Evaluating Transformer Language Models on Arithmetic Operations Using Number Decomposition 21 Apr 2023 · 1 repository · arXiv:2304.10977
-
Inducing anxiety in large language models can induce bias 21 Apr 2023 · 0 repositories · arXiv:2304.11111
-
The Dark Side of ChatGPT: Legal and Ethical Challenges from Stochastic Parrots and Hallucination 21 Apr 2023 · 0 repositories · arXiv:2304.14347
-
Who's the Best Detective? LLMs vs. MLs in Detecting Incoherent Fourth Grade Math Answers 21 Apr 2023 · 0 repositories · arXiv:2304.11257
-
Enhancing object detection robustness: A synthetic and natural perturbation approach 20 Apr 2023 · 0 repositories · arXiv:2304.10622
-
Meta Semantics: Towards better natural language understanding and reasoning 20 Apr 2023 · 0 repositories · arXiv:2304.10663
-
Safety Assessment of Chinese Large Language Models 20 Apr 2023 · 2 repositories · arXiv:2304.10436
-
SINC: Spatial Composition of 3D Human Motions for Simultaneous Action Generation 20 Apr 2023 · 0 repositories · arXiv:2304.10417
-
Catch Me If You Can: Identifying Fraudulent Physician Reviews with Large Language Models Using Generative Pre-Trained Transformers 19 Apr 2023 · 0 repositories · arXiv:2304.09948
-
GeneGPT: Augmenting Large Language Models with Domain Tools for Improved Access to Biomedical Information 19 Apr 2023 · 1 repository · arXiv:2304.09667Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
How to Do Things with Deep Learning Code 19 Apr 2023 · 0 repositories · arXiv:2304.09406
-
Supporting Human-AI Collaboration in Auditing LLMs with LLMs 19 Apr 2023 · 0 repositories · arXiv:2304.09991
-
SurgicalGPT: End-to-End Language-Vision GPT for Visual Question Answering in Surgery 19 Apr 2023 · 1 repository · arXiv:2304.09974Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
BIM-GPT: a Prompt-Based Virtual Assistant Framework for BIM Information Retrieval 18 Apr 2023 · 0 repositories · arXiv:2304.09333
-
CancerGPT: Few-shot Drug Pair Synergy Prediction using Large Pre-trained Language Models 18 Apr 2023 · 0 repositories · arXiv:2304.10946
-
LLM-based Interaction for Content Generation: A Case Study on the Perception of Employees in an IT department 18 Apr 2023 · 0 repositories · arXiv:2304.09064
-
An Empirical Study of Multitask Learning to Improve Open Domain Dialogue Systems 17 Apr 2023 · 1 repository · arXiv:2304.08115
-
From Zero to Hero: Examining the Power of Symbolic Tasks in Instruction Tuning 17 Apr 2023 · 1 repository · arXiv:2304.07995Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Supporting Qualitative Analysis with Large Language Models: Combining Codebook with GPT-3 for Deductive Coding 17 Apr 2023 · 0 repositories · arXiv:2304.10548
-
ArguGPT: evaluating, understanding and identifying argumentative essays generated by GPT models 16 Apr 2023 · 2 repositories · arXiv:2304.07666
-
Enhancing Automated Program Repair through Fine-tuning and Prompt Engineering 16 Apr 2023 · 0 repositories · arXiv:2304.07840
-
Sabiá: Portuguese Large Language Models 16 Apr 2023 · 0 repositories · arXiv:2304.07880
-
SikuGPT: A Generative Pre-trained Model for Intelligent Information Processing of Ancient Texts from the Perspective of Digital Humanities 16 Apr 2023 · 1 repository · arXiv:2304.07778
-
Towards Better Instruction Following Language Models for Chinese: Investigating the Impact of Training Data and Evaluation 16 Apr 2023 · 2 repositories · arXiv:2304.07854
-
Can ChatGPT Forecast Stock Price Movements? Return Predictability and Large Language Models 15 Apr 2023 · 0 repositories · arXiv:2304.07619
-
API-Bank: A Comprehensive Benchmark for Tool-Augmented LLMs 14 Apr 2023 · 2 repositories · arXiv:2304.08244Syntology 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples)
-
ChatGPT: Applications, Opportunities, and Threats 14 Apr 2023 · 0 repositories · arXiv:2304.09103
-
MedAlpaca -- An Open-Source Collection of Medical Conversational AI Models and Training Data 14 Apr 2023 · 1 repository · arXiv:2304.08247Syntology 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples) · 1 pointer-only (licence)
-
Stochastic Code Generation 14 Apr 2023 · 0 repositories · arXiv:2304.08243
-
ChatGPT cites the most-cited articles and journals, relying solely on Google Scholar's citation counts. As a result, AI may amplify the Matthew Effect in environmental science 13 Apr 2023 · 0 repositories · arXiv:2304.06794
-
PGTask: Introducing the Task of Profile Generation from Dialogues 13 Apr 2023 · 1 repository · arXiv:2304.06634
-
Shall We Pretrain Autoregressive Language Models with Retrieval? A Comprehensive Study 13 Apr 2023 · 1 repository · arXiv:2304.06762
-
What does CLIP know about a red circle? Visual prompt engineering for VLMs 13 Apr 2023 · 0 repositories · arXiv:2304.06712
-
Detection of Fake Generated Scientific Abstracts 12 Apr 2023 · 1 repository · arXiv:2304.06148
-
Evaluation of ChatGPT Model for Vulnerability Detection 12 Apr 2023 · 0 repositories · arXiv:2304.07232
-
Localizing Model Behavior with Path Patching 12 Apr 2023 · 1 repository · arXiv:2304.05969Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
Approximating Online Human Evaluation of Social Chatbots with Prompting 11 Apr 2023 · 0 repositories · arXiv:2304.05253
-
Bayesian Optimization of Catalysis With In-Context Learning 11 Apr 2023 · 2 repositories · arXiv:2304.05341
-
Distinguishing ChatGPT(-3.5, -4)-generated and human-written papers through Japanese stylometric analysis 11 Apr 2023 · 0 repositories · arXiv:2304.05534
-
Multi-step Jailbreaking Privacy Attacks on ChatGPT 11 Apr 2023 · 1 repository · arXiv:2304.05197Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples)
-
Training Large Language Models Efficiently with Sparsity and Dataflow 11 Apr 2023 · 0 repositories · arXiv:2304.05511
-
Automated Reading Passage Generation with OpenAI's Large Language Model 10 Apr 2023 · 0 repositories · arXiv:2304.04616
-
On the Possibilities of AI-Generated Text Detection 10 Apr 2023 · 0 repositories · arXiv:2304.04736
-
Are Large Language Models Ready for Healthcare? A Comparative Study on Clinical Language Understanding 9 Apr 2023 · 1 repository · arXiv:2304.05368
-
GPT4Rec: A Generative Framework for Personalized Recommendation and User Interests Interpretation 8 Apr 2023 · 0 repositories · arXiv:2304.03879
-
ChatGPT-Crawler: Find out if ChatGPT really knows what it's talking about 6 Apr 2023 · 0 repositories · arXiv:2304.03325
-
GPT detectors are biased against non-native English writers 6 Apr 2023 · 2 repositories · arXiv:2304.02819Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Making AI Less "Thirsty": Uncovering and Addressing the Secret Water Footprint of AI Models 6 Apr 2023 · 1 repository · arXiv:2304.03271Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
Towards Interpretable Mental Health Analysis with Large Language Models 6 Apr 2023 · 2 repositories · arXiv:2304.03347
-
Zero-Shot Next-Item Recommendation using Large Pretrained Language Models 6 Apr 2023 · 1 repository · arXiv:2304.03153Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Conceptual structure coheres in human cognition but not in large language models 5 Apr 2023 · 0 repositories · arXiv:2304.02754
-
Document-Level Machine Translation with Large Language Models 5 Apr 2023 · 1 repository · arXiv:2304.02210
-
Large Language Models as Master Key: Unlocking the Secrets of Materials Science with GPT 5 Apr 2023 · 0 repositories · arXiv:2304.02213
-
Blockwise Compression of Transformer-based Models without Retraining 4 Apr 2023 · 0 repositories · arXiv:2304.01483
-
Geotechnical Parrot Tales (GPT): Harnessing Large Language Models in geotechnical engineering 4 Apr 2023 · 0 repositories · arXiv:2304.02138
-
GPT-4 to GPT-3.5: 'Hold My Scalpel' -- A Look at the Competency of OpenAI's GPT on the Plastic Surgery In-Service Training Exam 4 Apr 2023 · 0 repositories · arXiv:2304.01503
-
Is ChatGPT a Highly Fluent Grammatical Error Correction System? A Comprehensive Evaluation 4 Apr 2023 · 0 repositories · arXiv:2304.01746
-
LLM-Adapters: An Adapter Family for Parameter-Efficient Fine-Tuning of Large Language Models 4 Apr 2023 · 2 repositories · arXiv:2304.01933Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
REFINER: Reasoning Feedback on Intermediate Representations 4 Apr 2023 · 1 repository · arXiv:2304.01904Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Summary of ChatGPT-Related Research and Perspective Towards the Future of Large Language Models 4 Apr 2023 · 0 repositories · arXiv:2304.01852
-
GreekBART: The First Pretrained Greek Sequence-to-Sequence Model 3 Apr 2023 · 2 repositories · arXiv:2304.00869
-
Does Human Collaboration Enhance the Accuracy of Identifying LLM-Generated Deepfake Texts? 3 Apr 2023 · 2 repositories · arXiv:2304.01002Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
LLMMaps -- A Visual Metaphor for Stratified Evaluation of Large Language Models 2 Apr 2023 · 1 repository · arXiv:2304.00457
-
Evaluating GPT-4 and ChatGPT on Japanese Medical Licensing Examinations 31 Mar 2023 · 1 repository · arXiv:2303.18027
-
GPT-4 can pass the Korean National Licensing Examination for Korean Medicine Doctors 31 Mar 2023 · 0 repositories · arXiv:2303.17807
-
Aligning a medium-size GPT model in English to a small closed domain in Spanish 30 Mar 2023 · 0 repositories · arXiv:2303.17649
-
Evaluation of GPT and BERT-based models on identifying protein-protein interactions in biomedical text 30 Mar 2023 · 0 repositories · arXiv:2303.17728
-
Humans in Humans Out: On GPT Converging Toward Common Sense in both Success and Failure 30 Mar 2023 · 0 repositories · arXiv:2303.17276
-
Synthesis of Mathematical programs from Natural Language Specifications 30 Mar 2023 · 0 repositories · arXiv:2304.03287
-
Advances in apparent conceptual physics reasoning in GPT-4 29 Mar 2023 · 0 repositories · arXiv:2303.17012
-
AnnoLLM: Making Large Language Models to Be Better Crowdsourced Annotators 29 Mar 2023 · 2 repositories · arXiv:2303.16854Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
AutoAD: Movie Description in Context 29 Mar 2023 · 1 repository · arXiv:2303.16899Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 8 unverified (of 13 harvested samples)
-
Evaluating GPT-3.5 and GPT-4 Models on Brazilian University Admission Exams 29 Mar 2023 · 1 repository · arXiv:2303.17003Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 6 pointer-only (licence)
-
How do decoding algorithms distribute information in dialogue responses? 29 Mar 2023 · 0 repositories · arXiv:2303.17006