Methods › General › Learning Rate Schedules › Cosine Annealing › Papers, page 29
Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,965 · with a code link: 1,734 · where Syntology ran a sample: 627 (513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (627 of 3,965 tagged: 513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument)
Page 29 of 40: papers 2,801 to 2,900 of 3,965, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Ten Quick Tips for Harnessing the Power of ChatGPT/GPT-4 in Computational Biology 29 Mar 2023 · 1 repository · arXiv:2303.16429
-
ViewRefer: Grasp the Multi-view Knowledge for 3D Visual Grounding with GPT and Prototype Guidance 29 Mar 2023 · 7 repositories · arXiv:2303.16894Syntology official (archive's flag): 4 ran · 8 ran (of which 2 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Improving Large Language Models for Clinical Named Entity Recognition via Prompt Engineering 29 Mar 2023 · 1 repository · arXiv:2303.16416
-
Explicit Planning Helps Language Models in Logical Reasoning 28 Mar 2023 · 2 repositories · arXiv:2303.15714Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
On Codex Prompt Engineering for OCL Generation: An Empirical Study 28 Mar 2023 · 0 repositories · arXiv:2303.16244
-
Zero-Shot Generalizable End-to-End Task-Oriented Dialog System using Context Summarization and Domain Schema 28 Mar 2023 · 1 repository · arXiv:2303.16252
-
KPEval: Towards Fine-Grained Semantic-Based Keyphrase Evaluation 27 Mar 2023 · 1 repository · arXiv:2303.15422
-
Analyzing the Performance of GPT-3.5 and GPT-4 in Grammatical Error Correction 25 Mar 2023 · 0 repositories · arXiv:2303.14342
-
Can Large Language Models assist in Hazard Analysis? 25 Mar 2023 · 0 repositories · arXiv:2303.15473
-
GPT is becoming a Turing machine: Here are some ways to program it 25 Mar 2023 · 0 repositories · arXiv:2303.14310
-
"Get ready for a party": Exploring smarter smart spaces with help from large language models 24 Mar 2023 · 1 repository · arXiv:2303.14143
-
Personalizing Task-oriented Dialog Systems via Zero-shot Generalizable Reward Function 24 Mar 2023 · 0 repositories · arXiv:2303.13797
-
SEAL: Semantic Frame Execution And Localization for Perceiving Afforded Robot Actions 24 Mar 2023 · 0 repositories · arXiv:2303.14067
-
GesGPT: Speech Gesture Synthesis With Text Parsing from ChatGPT 23 Mar 2023 · 0 repositories · arXiv:2303.13013
-
Generate labeled training data using Prompt Programming and GPT-3. An example of Big Five Personality Classification 22 Mar 2023 · 0 repositories · arXiv:2303.12279
-
A Complete Survey on Generative AI (AIGC): Is ChatGPT from GPT-4 to GPT-5 All You Need? 21 Mar 2023 · 0 repositories · arXiv:2303.11717
-
ChatGPT and a New Academic Reality: Artificial Intelligence-Written Research Papers and the Ethics of the Large Language Models in Scholarly Publishing 21 Mar 2023 · 0 repositories · arXiv:2303.13367
-
cTBLS: Augmenting Large Language Models with Conversational Tables 21 Mar 2023 · 1 repository · arXiv:2303.12024
-
Learning A Sparse Transformer Network for Effective Image Deraining 21 Mar 2023 · 1 repository · arXiv:2303.11950Syntology official (archive's flag): 8 ran · 8 ran (of which 8 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified; every one of the 8 samples that ran constructed an object rather than computing a result (of 11 harvested samples) · 11 pointer-only (licence)
-
Sparse-IFT: Sparse Iso-FLOP Transformations for Maximizing Training Efficiency 21 Mar 2023 · 2 repositories · arXiv:2303.11525Syntology official (archive's flag): 21 ran · 21 ran (of which 0 constructed an object rather than computing a result; 20 with no instrument failure: 0 honoured, 0 violated, 20 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 24 harvested samples) · 2 pointer-only (licence)
-
Capabilities of GPT-4 on Medical Challenge Problems 20 Mar 2023 · 1 repository · arXiv:2303.13375
-
Learning Behavior Recognition in Smart Classroom with Multiple Students Based on YOLOv5 20 Mar 2023 · 0 repositories · arXiv:2303.10916
-
Mind meets machine: Unravelling GPT-4's cognitive psychology 20 Mar 2023 · 0 repositories · arXiv:2303.11436
-
A Comprehensive Capability Analysis of GPT-3 and GPT-3.5 Series Models 18 Mar 2023 · 0 repositories · arXiv:2303.10420
-
SPDF: Sparse Pre-training and Dense Fine-tuning for Large Language Models 18 Mar 2023 · 0 repositories · arXiv:2303.10464
-
GPTs are GPTs: An Early Look at the Labor Market Impact Potential of Large Language Models 17 Mar 2023 · 0 repositories · arXiv:2303.10130
-
Block-wise Bit-Compression of Transformer-based Models 16 Mar 2023 · 0 repositories · arXiv:2303.09184
-
Can Generative Pre-trained Transformers (GPT) Pass Assessments in Higher Education Programming Courses? 16 Mar 2023 · 0 repositories · arXiv:2303.09325
-
Towards Commonsense Knowledge based Fuzzy Systems for Supporting Size-Related Fine-Grained Object Detection 16 Mar 2023 · 1 repository · arXiv:2303.09026
-
Jump to Conclusions: Short-Cutting Transformers With Linear Transformations 16 Mar 2023 · 2 repositories · arXiv:2303.09435Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Towards the Scalable Evaluation of Cooperativeness in Language Models 16 Mar 2023 · 0 repositories · arXiv:2303.13360
-
Automated Interactive Domain-Specific Conversational Agents that Understand Human Dialogs 15 Mar 2023 · 0 repositories · arXiv:2303.08941
-
GCRE-GPT: A Generative Model for Comparative Relation Extraction 15 Mar 2023 · 0 repositories · arXiv:2303.08601
-
SelfCheckGPT: Zero-Resource Black-Box Hallucination Detection for Generative Large Language Models 15 Mar 2023 · 1 repository · arXiv:2303.08896Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
Can ChatGPT Replace Traditional KBQA Models? An In-depth Analysis of the Question Answering Performance of the GPT LLM Family 14 Mar 2023 · 2 repositories · arXiv:2303.07992
-
RE-MOVE: An Adaptive Policy Design for Robotic Navigation Tasks in Dynamic Environments via Language-Based Feedback 14 Mar 2023 · 0 repositories · arXiv:2303.07622
-
Large Language Models in the Workplace: A Case Study on Prompt Engineering for Job Type Classification 13 Mar 2023 · 0 repositories · arXiv:2303.07142
-
Transformer-based World Models Are Happy With 100k Interactions 13 Mar 2023 · 1 repository · arXiv:2303.07109Syntology official (archive's flag): 16 ran · 16 ran (of which 6 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 1 violated, 6 with no contract checked; 8 where Syntology's instrument failed) · 9 unverified (of 25 harvested samples)
-
Large Language Models Know Your Contextual Search Intent: A Prompting Framework for Conversational Search 12 Mar 2023 · 2 repositories · arXiv:2303.06573
-
Learning Combinatorial Prompts for Universal Controllable Image Captioning 11 Mar 2023 · 0 repositories · arXiv:2303.06338
-
Algorithmic Ghost in the Research Shell: Large Language Models and Academic Knowledge Creation in Management Research 10 Mar 2023 · 0 repositories · arXiv:2303.07304
-
ChatGPT may Pass the Bar Exam soon, but has a Long Way to Go for the LexGLUE benchmark 9 Mar 2023 · 1 repository · arXiv:2304.12202
-
ICL-D3IE: In-Context Learning with Diverse Demonstrations Updating for Document Information Extraction 9 Mar 2023 · 1 repository · arXiv:2303.05063
-
Large Language Models (GPT) Struggle to Answer Multiple-Choice Questions about Code 9 Mar 2023 · 0 repositories · arXiv:2303.08033
-
ChatGPT Participates in a Computer Science Exam 8 Mar 2023 · 1 repository · arXiv:2303.09461
-
Cost-Effective Hyperparameter Optimization for Large Language Model Generation Inference 8 Mar 2023 · 3 repositories · arXiv:2303.04673Syntology community repositories only · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Stealing the Decoding Algorithms of Language Models 8 Mar 2023 · 1 repository · arXiv:2303.04729
-
A Comprehensive Survey of AI-Generated Content (AIGC): A History of Generative AI from GAN to ChatGPT 7 Mar 2023 · 1 repository · arXiv:2303.04226
-
Towards Zero-Shot Functional Compositionality of Language Models 6 Mar 2023 · 1 repository · arXiv:2303.03103
-
Industry Risk Assessment via Hierarchical Financial Data Using Stock Market Sentiment Indicators 5 Mar 2023 · 0 repositories · arXiv:2303.02707
-
Prompt, Generate, then Cache: Cascade of Foundation Models makes Strong Few-shot Learners 3 Mar 2023 · 3 repositories · arXiv:2303.02151Syntology official: not harvested · 0 ran · 1 unverified (of 1 harvested sample)
-
Prophet: Prompting Large Language Models with Complementary Answer Heuristics for Knowledge-based Visual Question Answering 3 Mar 2023 · 1 repository · arXiv:2303.01903Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
WiCE: Real-World Entailment for Claims in Wikipedia 2 Mar 2023 · 2 repositories · arXiv:2303.01432Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
A Framework for Neurosymbolic Robot Action Planning using Large Language Models 1 Mar 2023 · 1 repository · arXiv:2303.00438Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
How Robust is GPT-3.5 to Predecessors? A Comprehensive Study on Language Understanding Tasks 1 Mar 2023 · 0 repositories · arXiv:2303.00293
-
ToxVis: Enabling Interpretability of Implicit vs. Explicit Toxicity Detection Models with Interactive Visualization 1 Mar 2023 · 0 repositories · arXiv:2303.09402
-
Zero-Shot Cross-Lingual Summarization via Large Language Models 28 Feb 2023 · 0 repositories · arXiv:2302.14229
-
Information-Restricted Neural Language Models Reveal Different Brain Regions' Sensitivity to Semantics, Syntax and Context 28 Feb 2023 · 1 repository · arXiv:2302.14389Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Large Language Models Are State-of-the-Art Evaluators of Translation Quality 28 Feb 2023 · 4 repositories · arXiv:2302.14520
-
Sampled Transformer for Point Sets 28 Feb 2023 · 0 repositories · arXiv:2302.14346
-
Inseq: An Interpretability Toolkit for Sequence Generation Models 27 Feb 2023 · 2 repositories · arXiv:2302.13942
-
LLaMA: Open and Efficient Foundation Language Models 27 Feb 2023 · 57 repositories · arXiv:2302.13971Syntology official: no sample here; runs from other or unrecorded repositories · 37 ran (of which 9 constructed an object rather than computing a result; 25 with no instrument failure: 3 honoured, 0 violated, 22 with no contract checked; 12 where Syntology's instrument failed) · 21 unverified (of 58 harvested samples) · 4 pointer-only (licence)
-
Reward Design with Language Models 27 Feb 2023 · 1 repository · arXiv:2303.00001
-
Supervised Virtual-to-Real Domain Adaptation for Object Detection Task using YOLO 27 Feb 2023 · 0 repositories · arXiv:2302.13891
-
Systematic Rectification of Language Models via Dead-end Analysis 27 Feb 2023 · 1 repository · arXiv:2302.14003Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
Comparing Sentence-Level Suggestions to Message-Level Suggestions in AI-Mediated Communication 26 Feb 2023 · 0 repositories · arXiv:2302.13382
-
Fast Attention Requires Bounded Entries 26 Feb 2023 · 0 repositories · arXiv:2302.13214
-
Human-in-the-Loop Schema Induction 25 Feb 2023 · 0 repositories · arXiv:2302.13048
-
Spanish Built Factual Freectianary (Spanish-BFF): the first AI-generated free dictionary 24 Feb 2023 · 0 repositories · arXiv:2302.12746
-
Testing AI on language comprehension tasks reveals insensitivity to underlying meaning 23 Feb 2023 · 0 repositories · arXiv:2302.12313
-
What makes a language easy to deep-learn? Deep neural networks and humans similarly benefit from compositional structure 23 Feb 2023 · 1 repository · arXiv:2302.12239Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
kNN-Adapter: Efficient Domain Adaptation for Black-Box Language Models 21 Feb 2023 · 0 repositories · arXiv:2302.10879
-
Large-scale Multi-Modal Pre-trained Models: A Comprehensive Survey 20 Feb 2023 · 1 repository · arXiv:2302.10035
-
ChatIE: Zero-Shot Information Extraction via Chatting with ChatGPT 20 Feb 2023 · 1 repository · arXiv:2302.10205
-
Table Tennis Stroke Detection and Recognition Using Ball Trajectory Data 19 Feb 2023 · 0 repositories · arXiv:2302.09657
-
A Comprehensive Survey on Pretrained Foundation Models: A History from BERT to ChatGPT 18 Feb 2023 · 0 repositories · arXiv:2302.09419
-
How Good Are GPT Models at Machine Translation? A Comprehensive Evaluation 18 Feb 2023 · 1 repository · arXiv:2302.09210Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Bounding the Capabilities of Large Language Models in Open Text Generation with Prompt Constraints 17 Feb 2023 · 1 repository · arXiv:2302.09185
-
Conveying the Predicted Future to Users: A Case Study of Story Plot Prediction 17 Feb 2023 · 1 repository · arXiv:2302.09122
-
GPT4MIA: Utilizing Generative Pre-trained Transformer (GPT-3) as A Plug-and-Play Transductive Model for Medical Image Analysis 17 Feb 2023 · 0 repositories · arXiv:2302.08722
-
PAC Prediction Sets for Large Language Models of Code 17 Feb 2023 · 1 repository · arXiv:2302.08703
-
Prompting Large Language Models With the Socratic Method 17 Feb 2023 · 0 repositories · arXiv:2303.08769
-
Foundation Models for Natural Language Processing -- Pre-trained Language Models Integrating Media 16 Feb 2023 · 0 repositories · arXiv:2302.08575
-
For Generated Text, Is NLI-Neutral Text the Best Text? 16 Feb 2023 · 1 repository · arXiv:2302.08577
-
Commonsense Reasoning for Conversational AI: A Survey of the State of the Art 15 Feb 2023 · 0 repositories · arXiv:2302.07926
-
Learning Performance-Improving Code Edits 15 Feb 2023 · 2 repositories · arXiv:2302.07867Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 5 where Syntology's instrument failed) · 11 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Tree-Based Representation and Generation of Natural and Mathematical Language 15 Feb 2023 · 1 repository · arXiv:2302.07974
-
ScatterShot: Interactive In-context Example Curation for Text Transformation 14 Feb 2023 · 1 repository · arXiv:2302.07346
-
Diminished Diversity-of-Thought in a Standard Large Language Model 13 Feb 2023 · 0 repositories · arXiv:2302.07267
-
Can GPT-3 Perform Statutory Reasoning? 13 Feb 2023 · 1 repository · arXiv:2302.06100
-
Detection and Segmentation of Pancreas using Morphological Snakes and Deep Convolutional Neural Networks 13 Feb 2023 · 0 repositories · arXiv:2302.06356
-
STREET: A Multi-Task Structured Reasoning and Explanation Benchmark 13 Feb 2023 · 0 repositories · arXiv:2302.06729
-
Academic Writing with GPT-3.5: Reflections on Practices, Efficacy and Transparency 12 Feb 2023 · 0 repositories · arXiv:2304.11079
-
A Brief Report on LawGPT 1.0: A Virtual Legal Assistant Based on GPT-3 11 Feb 2023 · 0 repositories · arXiv:2302.05729
-
Combat AI With AI: Counteract Machine-Generated Fake Restaurant Reviews on Social Media 10 Feb 2023 · 1 repository · arXiv:2302.07731
-
FairPy: A Toolkit for Evaluation of Prediction Biases and their Mitigation in Large Language Models 10 Feb 2023 · 1 repository · arXiv:2302.05508
-
GTR-CTRL: Instrument and Genre Conditioning for Guitar-Focused Music Generation with Transformers 10 Feb 2023 · 0 repositories · arXiv:2302.05393
-
The Wisdom of Hindsight Makes Language Models Better Instruction Followers 10 Feb 2023 · 1 repository · arXiv:2302.05206
-
Translating Natural Language to Planning Goals with Large-Language Models 10 Feb 2023 · 1 repository · arXiv:2302.05128
-
Better by you, better than me, chatgpt3 as writing assistance in students essays 9 Feb 2023 · 0 repositories · arXiv:2302.04536