Methods › General › Learning Rate Schedules › Cosine Annealing › Papers, page 21
Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,965 · with a code link: 1,734 · where Syntology ran a sample: 627 (513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (627 of 3,965 tagged: 513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument)
Page 21 of 40: papers 2,001 to 2,100 of 3,965, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
RCAgent: Cloud Root Cause Analysis by Autonomous Agents with Tool-Augmented Large Language Models 25 Oct 2023 · 0 repositories · arXiv:2310.16340
-
A Communication Theory Perspective on Prompting Engineering Methods for Large Language Models 24 Oct 2023 · 0 repositories · arXiv:2310.18358
-
A Language Model with Limited Memory Capacity Captures Interference in Human Sentence Processing 24 Oct 2023 · 0 repositories · arXiv:2310.16142
-
AI-enhanced Auto-correction of Programming Exercises: How Effective is GPT-3.5? 24 Oct 2023 · 0 repositories · arXiv:2311.10737
-
Background Summarization of Event Timelines 24 Oct 2023 · 1 repository · arXiv:2310.16197
-
Dissecting In-Context Learning of Translations in GPTs 24 Oct 2023 · 0 repositories · arXiv:2310.15987
-
Fighting Fire with Fire: The Dual Role of LLMs in Crafting and Detecting Elusive Disinformation 24 Oct 2023 · 1 repository · arXiv:2310.15515
-
Learning From Free-Text Human Feedback -- Collect New Datasets Or Extend Existing Ones? 24 Oct 2023 · 1 repository · arXiv:2310.15758Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
The Janus Interface: How Fine-Tuning in Large Language Models Amplifies the Privacy Risks 24 Oct 2023 · 1 repository · arXiv:2310.15469
-
TRAMS: Training-free Memory Selection for Long-range Language Modeling 24 Oct 2023 · 1 repository · arXiv:2310.15494Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
What Makes it Ok to Set a Fire? Iterative Self-distillation of Contexts and Rationales for Disambiguating Defeasible Social and Moral Situations 24 Oct 2023 · 0 repositories · arXiv:2310.15431
-
Causal Inference Using LLM-Guided Discovery 23 Oct 2023 · 0 repositories · arXiv:2310.15117
-
Evaluating Spatial Understanding of Large Language Models 23 Oct 2023 · 1 repository · arXiv:2310.14540Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples)
-
Evaluating the Knowledge Base Completion Potential of GPT 23 Oct 2023 · 0 repositories · arXiv:2310.14771
-
Generative Pre-trained Transformer for Vietnamese Community-based COVID-19 Question Answering 23 Oct 2023 · 0 repositories · arXiv:2310.14602
-
GPT-4 as an Effective Zero-Shot Evaluator for Scientific Figure Captions 23 Oct 2023 · 0 repositories · arXiv:2310.15405
-
InstructExcel: A Benchmark for Natural Language Instruction in Excel 23 Oct 2023 · 0 repositories · arXiv:2310.14495
-
Language Models Hallucinate, but May Excel at Fact Verification 23 Oct 2023 · 1 repository · arXiv:2310.14564Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
LINC: A Neurosymbolic Approach for Logical Reasoning by Combining Language Models with First-Order Logic Provers 23 Oct 2023 · 1 repository · arXiv:2310.15164Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
LLM-in-the-loop: Leveraging Large Language Model for Thematic Analysis 23 Oct 2023 · 1 repository · arXiv:2310.15100Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Prefix-Tuning Based Unsupervised Text Style Transfer 23 Oct 2023 · 0 repositories · arXiv:2310.14599
-
TeleQnA: A Benchmark Dataset to Assess Large Language Models Telecommunications Knowledge 23 Oct 2023 · 1 repository · arXiv:2310.15051
-
Establishing Vocabulary Tests as a Benchmark for Evaluating Large Language Models 23 Oct 2023 · 1 repository · arXiv:2310.14703
-
Towards a Mechanistic Interpretation of Multi-Step Reasoning Capabilities of Language Models 23 Oct 2023 · 2 repositories · arXiv:2310.14491Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Can Language Models Laugh at YouTube Short-form Videos? 22 Oct 2023 · 1 repository · arXiv:2310.14159
-
Is ChatGPT a game changer for geocoding -- a benchmark for geocoding address parsing techniques 22 Oct 2023 · 1 repository · arXiv:2310.14360
-
Text generation for dataset augmentation in security classification tasks 22 Oct 2023 · 1 repository · arXiv:2310.14429
-
GEMBA-MQM: Detecting Translation Quality Error Spans with GPT-4 21 Oct 2023 · 1 repository · arXiv:2310.13988Syntology 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
HateRephrase: Zero- and Few-Shot Reduction of Hate Intensity in Online Posts using Large Language Models 21 Oct 2023 · 0 repositories · arXiv:2310.13985
-
A Simple Baseline for Knowledge-Based Visual Question Answering 20 Oct 2023 · 0 repositories · arXiv:2310.13570Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
AllTogether: Investigating the Efficacy of Spliced Prompt for Web Navigation using Large Language Models 20 Oct 2023 · 0 repositories · arXiv:2310.18331
-
Application of deep learning for livestock behaviour recognition: A systematic literature review 20 Oct 2023 · 0 repositories · arXiv:2310.13483
-
Cache me if you Can: an Online Cost-aware Teacher-Student framework to Reduce the Calls to Large Language Models 20 Oct 2023 · 1 repository · arXiv:2310.13395Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Challenges and Contributing Factors in the Utilization of Large Language Models (LLMs) 20 Oct 2023 · 0 repositories · arXiv:2310.13343
-
She had Cobalt Blue Eyes: Prompt Testing to Create Aligned and Sustainable Language Models 20 Oct 2023 · 0 repositories · arXiv:2310.18333
-
Equivariant Transformer is all you need 20 Oct 2023 · 0 repositories · arXiv:2310.13222
-
Foundation Model's Embedded Representations May Detect Distribution Shift 20 Oct 2023 · 0 repositories · arXiv:2310.13836
-
Robust Training for Conversational Question Answering Models with Reinforced Reformulation Generation 20 Oct 2023 · 0 repositories · arXiv:2310.13505
-
The Perils & Promises of Fact-checking with Large Language Models 20 Oct 2023 · 0 repositories · arXiv:2310.13549
-
WordArt Designer: User-Driven Artistic Typography Synthesis using Large Language Models 20 Oct 2023 · 0 repositories · arXiv:2310.18332
-
AgentTuning: Enabling Generalized Agent Abilities for LLMs 19 Oct 2023 · 1 repository · arXiv:2310.12823Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Exploring In-Context Learning of Textless Speech Language Model for Speech Classification Tasks 19 Oct 2023 · 0 repositories · arXiv:2310.12477
-
Experimental Narratives: A Comparison of Human Crowdsourced Storytelling and AI Storytelling 19 Oct 2023 · 0 repositories · arXiv:2310.12902
-
Identifying and Adapting Transformer-Components Responsible for Gender Bias in an English Language Model 19 Oct 2023 · 1 repository · arXiv:2310.12611Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
LASER: Linear Compression in Wireless Distributed Optimization 19 Oct 2023 · 0 repositories · arXiv:2310.13033
-
Not All Countries Celebrate Thanksgiving: On the Cultural Dominance in Large Language Models 19 Oct 2023 · 0 repositories · arXiv:2310.12481
-
ExtractGPT: Exploring the Potential of Large Language Models for Product Attribute Value Extraction 19 Oct 2023 · 1 repository · arXiv:2310.12537
-
The Shifted and The Overlooked: A Task-oriented Investigation of User-GPT Interactions 19 Oct 2023 · 1 repository · arXiv:2310.12418Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Evaluating the Symbol Binding Ability of Large Language Models for Multiple-Choice Questions in Vietnamese General Education 18 Oct 2023 · 0 repositories · arXiv:2310.12059
-
Solving the multiplication problem of a large language model system using a graph-based method 18 Oct 2023 · 0 repositories · arXiv:2310.13016
-
Emergent AI-Assisted Discourse: Case Study of a Second Language Writer Authoring with ChatGPT 17 Oct 2023 · 0 repositories · arXiv:2310.10903
-
LLMs as Hackers: Autonomous Linux Privilege Escalation Attacks 17 Oct 2023 · 1 repository · arXiv:2310.11409
-
Intent Detection and Slot Filling for Home Assistants: Dataset and Analysis for Bangla and Sylheti 17 Oct 2023 · 1 repository · arXiv:2310.10935
-
Probing the Creativity of Large Language Models: Can models produce divergent semantic association? 17 Oct 2023 · 1 repository · arXiv:2310.11158Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Utilising a Large Language Model to Annotate Subject Metadata: A Case Study in an Australian National Research Data Catalogue 17 Oct 2023 · 0 repositories · arXiv:2310.11318
-
Approximating Two-Layer Feedforward Networks for Efficient Transformers 16 Oct 2023 · 2 repositories · arXiv:2310.10837Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Battle of the Large Language Models: Dolly vs LLaMA vs Vicuna vs Guanaco vs Bard vs ChatGPT -- A Text-to-SQL Parsing Comparison 16 Oct 2023 · 0 repositories · arXiv:2310.10190
-
BioPlanner: Automatic Evaluation of LLMs on Protocol Planning in Biology 16 Oct 2023 · 1 repository · arXiv:2310.10632Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Data Contamination Through the Lens of Time 16 Oct 2023 · 1 repository · arXiv:2310.10628Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
Fine-tuning ChatGPT for Automatic Scoring 16 Oct 2023 · 0 repositories · arXiv:2310.10072
-
MoConVQ: Unified Physics-Based Motion Control via Scalable Discrete Representations 16 Oct 2023 · 0 repositories · arXiv:2310.10198
-
Prediction of Arabic Legal Rulings using Large Language Models 16 Oct 2023 · 0 repositories · arXiv:2310.10260
-
TRANSOM: An Efficient Fault-Tolerant System for Training LLMs 16 Oct 2023 · 1 repository · arXiv:2310.10046Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples)
-
Configuration Validation with Large Language Models 15 Oct 2023 · 0 repositories · arXiv:2310.09690
-
Image Augmentation with Controlled Diffusion for Weakly-Supervised Semantic Segmentation 15 Oct 2023 · 0 repositories · arXiv:2310.09760
-
Large Language Model-Aware In-Context Learning for Code Generation 15 Oct 2023 · 0 repositories · arXiv:2310.09748
-
Large Language Models for In-Context Student Modeling: Synthesizing Student's Behavior in Visual Programming 15 Oct 2023 · 1 repository · arXiv:2310.10690
-
Efficient Model-Agnostic Multi-Group Equivariant Networks 14 Oct 2023 · 0 repositories · arXiv:2310.09675
-
Assessing and Enhancing the Robustness of Large Language Models with Task Structure Variations for Logical Reasoning 13 Oct 2023 · 1 repository · arXiv:2310.09430Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples) · 1 pointer-only (licence)
-
From Words and Exercises to Wellness: Farsi Chatbot for Self-Attachment Technique 13 Oct 2023 · 0 repositories · arXiv:2310.09362
-
Human-in-the-loop Machine Translation with Large Language Model 13 Oct 2023 · 1 repository · arXiv:2310.08908
-
Table-GPT: Table-tuned GPT for Diverse Table Tasks 13 Oct 2023 · 0 repositories · arXiv:2310.09263
-
QUIK: Towards End-to-End 4-Bit Inference on Generative Large Language Models 13 Oct 2023 · 1 repository · arXiv:2310.09259Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples)
-
Jailbreaking Black Box Large Language Models in Twenty Queries 12 Oct 2023 · 1 repository · arXiv:2310.08419Syntology official (archive's flag): 2 ran · 2 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 7 harvested samples)
-
Large language models can replicate cross-cultural differences in personality 12 Oct 2023 · 0 repositories · arXiv:2310.10679
-
Multiclass Classification of Policy Documents with Large Language Models 12 Oct 2023 · 0 repositories · arXiv:2310.08167
-
Promptor: A Conversational and Autonomous Prompt Generation Agent for Intelligent Text Entry Techniques 12 Oct 2023 · 0 repositories · arXiv:2310.08101
-
QASiNa: Religious Domain Question Answering using Sirah Nabawiyah 12 Oct 2023 · 1 repository · arXiv:2310.08102
-
Training Generative Question-Answering on Synthetic Data Obtained from an Instruct-tuned Model 12 Oct 2023 · 0 repositories · arXiv:2310.08072
-
Diversity of Thought Improves Reasoning Abilities of LLMs 11 Oct 2023 · 0 repositories · arXiv:2310.07088
-
Do Large Language Models have Shared Weaknesses in Medical Question Answering? 11 Oct 2023 · 0 repositories · arXiv:2310.07225
-
Found in the Middle: Permutation Self-Consistency Improves Listwise Ranking in Large Language Models 11 Oct 2023 · 1 repository · arXiv:2310.07712Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
InstructRetro: Instruction Tuning post Retrieval-Augmented Pretraining 11 Oct 2023 · 1 repository · arXiv:2310.07713
-
Large Language Models Are Zero-Shot Time Series Forecasters 11 Oct 2023 · 2 repositories · arXiv:2310.07820Syntology official (archive's flag): 1 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 1 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Uncovering Hidden Connections: Iterative Search and Reasoning for Video-grounded Dialog 11 Oct 2023 · 2 repositories · arXiv:2310.07259Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Diffusion Models for Wireless Communications 11 Oct 2023 · 0 repositories · arXiv:2310.07312
-
Automated clinical coding using off-the-shelf large language models 10 Oct 2023 · 0 repositories · arXiv:2310.06552
-
GeoLLM: Extracting Geospatial Knowledge from Large Language Models 10 Oct 2023 · 1 repository · arXiv:2310.06213Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
GPT-4 as an Agronomist Assistant? Answering Agriculture Exams Using Large Language Models 10 Oct 2023 · 0 repositories · arXiv:2310.06225
-
Humans and language models diverge when predicting repeating text 10 Oct 2023 · 1 repository · arXiv:2310.06408Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Large Language Models for Propaganda Detection 10 Oct 2023 · 2 repositories · arXiv:2310.06422
-
LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression 10 Oct 2023 · 3 repositories · arXiv:2310.06839
-
SEER : A Knapsack approach to Exemplar Selection for In-Context HybridQA 10 Oct 2023 · 1 repository · arXiv:2310.06675Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 10 unverified (of 16 harvested samples)
-
Cabbage Sweeter than Cake? Analysing the Potential of Large Language Models for Learning Conceptual Spaces 9 Oct 2023 · 0 repositories · arXiv:2310.05481
-
Foundation Models Meet Visualizations: Challenges and Opportunities 9 Oct 2023 · 0 repositories · arXiv:2310.05771
-
Exploring the Maze of Multilingual Modeling 9 Oct 2023 · 0 repositories · arXiv:2310.05404
-
SC-Safety: A Multi-round Open-ended Question Adversarial Safety Benchmark for Large Language Models in Chinese 9 Oct 2023 · 0 repositories · arXiv:2310.05818
-
The Program Testing Ability of Large Language Models for Code 9 Oct 2023 · 0 repositories · arXiv:2310.05727
-
Are Emily and Greg Still More Employable than Lakisha and Jamal? Investigating Algorithmic Hiring Bias in the Era of ChatGPT 8 Oct 2023 · 0 repositories · arXiv:2310.05135
-
Distantly-Supervised Joint Extraction with Noise-Robust Learning 8 Oct 2023 · 1 repository · arXiv:2310.04994