Methods › General › Learning Rate Schedules › Cosine Annealing › Papers, page 17
Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,965 · with a code link: 1,734 · where Syntology ran a sample: 627 (513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (627 of 3,965 tagged: 513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument)
Page 17 of 40: papers 1,601 to 1,700 of 3,965, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Named Entity Recognition for Address Extraction in Speech-to-Text Transcriptions Using Synthetic Data 8 Feb 2024 · 0 repositories · arXiv:2402.05545
-
Zero-Shot Chain-of-Thought Reasoning Guided by Evolutionary Algorithms in Large Language Models 8 Feb 2024 · 0 repositories · arXiv:2402.05376
-
A Hypothesis-Driven Framework for the Analysis of Self-Rationalising Models 7 Feb 2024 · 1 repository · arXiv:2402.04787
-
Amortized Planning with Large-Scale Transformers: A Case Study on Chess 7 Feb 2024 · 1 repository · arXiv:2402.04494Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Improving Cross-Domain Low-Resource Text Generation through LLM Post-Editing: A Programmer-Interpreter Approach 7 Feb 2024 · 0 repositories · arXiv:2402.04609
-
Long Is More for Alignment: A Simple but Tough-to-Beat Baseline for Instruction Fine-Tuning 7 Feb 2024 · 1 repository · arXiv:2402.04833Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Advancing Legal Reasoning: The Integration of AI to Navigate Complexities and Biases in Global Jurisprudence with Semi-Automated Arbitration Processes (SAAPs) 6 Feb 2024 · 0 repositories · arXiv:2402.04140
-
Behind the Screen: Investigating ChatGPT's Dark Personality Traits and Conspiracy Beliefs 6 Feb 2024 · 0 repositories · arXiv:2402.04110
-
CEHR-GPT: Generating Electronic Health Records with Chronological Patient Timelines 6 Feb 2024 · 0 repositories · arXiv:2402.04400
-
Detecting Mode Collapse in Language Models via Narration 6 Feb 2024 · 0 repositories · arXiv:2402.04477
-
Large Language Models as an Indirect Reasoner: Contrapositive and Contradiction for Automated Reasoning 6 Feb 2024 · 0 repositories · arXiv:2402.03667
-
Large Language Models As MOOCs Graders 6 Feb 2024 · 0 repositories · arXiv:2402.03776
-
Leak, Cheat, Repeat: Data Contamination and Evaluation Malpractices in Closed-Source LLMs 6 Feb 2024 · 0 repositories · arXiv:2402.03927
-
Are Machines Better at Complex Reasoning? Unveiling Human-Machine Inference Gaps in Entailment Verification 6 Feb 2024 · 0 repositories · arXiv:2402.03686
-
Pard: Permutation-Invariant Autoregressive Diffusion for Graph Generation 6 Feb 2024 · 1 repository · arXiv:2402.03687
-
The Hedgehog & the Porcupine: Expressive Linear Attentions with Softmax Mimicry 6 Feb 2024 · 1 repository · arXiv:2402.04347
-
Training Language Models to Generate Text with Citations via Fine-grained Rewards 6 Feb 2024 · 1 repository · arXiv:2402.04315Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples) · 1 pointer-only (licence)
-
Reconstruct Your Previous Conversations! Comprehensively Investigating Privacy Leakage Risks in Conversations with GPT Models 5 Feb 2024 · 1 repository · arXiv:2402.02987
-
Harnessing PubMed User Query Logs for Post Hoc Explanations of Recommended Similar Articles 5 Feb 2024 · 0 repositories · arXiv:2402.03484
-
LLM Agents in Interaction: Measuring Personality Consistency and Linguistic Alignment in Interacting Populations of Large Language Models 5 Feb 2024 · 1 repository · arXiv:2402.02896
-
SWAG: Storytelling With Action Guidance 5 Feb 2024 · 1 repository · arXiv:2402.03483
-
UniMem: Towards a Unified View of Long-Context Large Language Models 5 Feb 2024 · 1 repository · arXiv:2402.03009
-
A Graph is Worth K Words: Euclideanizing Graph using Pure Transformer 4 Feb 2024 · 1 repository · arXiv:2402.02464Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
AutoTimes: Autoregressive Time Series Forecasters via Large Language Models 4 Feb 2024 · 1 repository · arXiv:2402.02370
-
GeReA: Question-Aware Prompt Captions for Knowledge-based Visual Question Answering 4 Feb 2024 · 1 repository · arXiv:2402.02503Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 18 harvested samples) · 18 pointer-only (licence)
-
Improving Assessment of Tutoring Practices using Retrieval-Augmented Generation 4 Feb 2024 · 0 repositories · arXiv:2402.14594
-
EffiBench: Benchmarking the Efficiency of Automatically Generated Code 3 Feb 2024 · 1 repository · arXiv:2402.02037Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Hierarchical Structure Enhances the Convergence and Generalizability of Linear Molecular Representation 3 Feb 2024 · 1 repository · arXiv:2402.02164
-
COMET: Generating Commit Messages using Delta Graph Context Representation 2 Feb 2024 · 0 repositories · arXiv:2402.01841
-
Can LLMs perform structured graph reasoning? 2 Feb 2024 · 1 repository · arXiv:2402.01805
-
Improving Sequential Recommendations with LLMs 2 Feb 2024 · 1 repository · arXiv:2402.01339Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Generation, Distillation and Evaluation of Motivational Interviewing-Style Reflections with a Foundational Language Model 1 Feb 2024 · 0 repositories · arXiv:2402.01051
-
Learning Planning-based Reasoning by Trajectories Collection and Process Reward Synthesizing 1 Feb 2024 · 0 repositories · arXiv:2402.00658
-
Self-Supervised Contrastive Pre-Training for Multivariate Point Processes 1 Feb 2024 · 0 repositories · arXiv:2402.00987
-
SPARQL Generation with Entity Pre-trained GPT for KG Question Answering 1 Feb 2024 · 1 repository · arXiv:2402.00969
-
Tiny Titans: Can Smaller Large Language Models Punch Above Their Weight in the Real World for Meeting Summarization? 1 Feb 2024 · 0 repositories · arXiv:2402.00841
-
Human-mediated Large Language Models for Robotic Intervention in Children with Autism Spectrum Disorders 1 Feb 2024 · 0 repositories · arXiv:2402.00260
-
ConSmax: Hardware-Friendly Alternative Softmax with Learnable Parameters 31 Jan 2024 · 1 repository · arXiv:2402.10930
-
Global-Liar: Factuality of LLMs over Time and Geographic Regions 31 Jan 2024 · 0 repositories · arXiv:2401.17839
-
Making a Long Story Short in Conversation Modeling 31 Jan 2024 · 0 repositories · arXiv:2402.00143
-
Mitigating the Influence of Distractor Tasks in LMs with Prior-Aware Decoding 31 Jan 2024 · 0 repositories · arXiv:2401.17692
-
Paramanu: A Family of Novel Efficient Generative Foundation Language Models for Indian Languages 31 Jan 2024 · 0 repositories · arXiv:2401.18034
-
Real Sparks of Artificial Intelligence and the Importance of Inner Interpretability 31 Jan 2024 · 0 repositories · arXiv:2402.00901
-
Uncertainty-Aware Explainable Recommendation with Large Language Models 31 Jan 2024 · 0 repositories · arXiv:2402.03366
-
A Preliminary Study on Using Large Language Models in Software Pentesting 30 Jan 2024 · 0 repositories · arXiv:2401.17459
-
LLaMP: Large Language Model Made Powerful for High-fidelity Materials Knowledge Retrieval and Distillation 30 Jan 2024 · 1 repository · arXiv:2401.17244Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
MT-Eval: A Multi-Turn Capabilities Evaluation Benchmark for Large Language Models 30 Jan 2024 · 1 repository · arXiv:2401.16745Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 2 pointer-only (licence)
-
Diverse, but Divisive: LLMs Can Exaggerate Gender Differences in Opinion Related to Harms of Misinformation 29 Jan 2024 · 0 repositories · arXiv:2401.16558
-
E-EVAL: A Comprehensive Chinese K-12 Education Evaluation Benchmark for Large Language Models 29 Jan 2024 · 1 repository · arXiv:2401.15927Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Leveraging Professional Radiologists' Expertise to Enhance LLMs' Evaluation for Radiology Reports 29 Jan 2024 · 0 repositories · arXiv:2401.16578
-
LLM4Vuln: A Unified Evaluation Framework for Decoupling and Enhancing LLMs' Vulnerability Reasoning 29 Jan 2024 · 0 repositories · arXiv:2401.16185
-
ReGAL: Refactoring Programs to Discover Generalizable Abstractions 29 Jan 2024 · 1 repository · arXiv:2401.16467Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
An Insight into Security Code Review with LLMs: Capabilities, Obstacles, and Influential Factors 29 Jan 2024 · 0 repositories · arXiv:2401.16310
-
TrackGPT -- A generative pre-trained transformer for cross-domain entity trajectory forecasting 29 Jan 2024 · 0 repositories · arXiv:2402.00066
-
ConvoSense: Overcoming Monotonous Commonsense Inferences for Conversational AI 27 Jan 2024 · 1 repository · arXiv:2401.15471
-
Enhancing Large Language Model Performance To Answer Questions and Extract Information More Accurately 27 Jan 2024 · 0 repositories · arXiv:2402.01722
-
Equipping Language Models with Tool Use Capability for Tabular Data Analysis in Finance 27 Jan 2024 · 0 repositories · arXiv:2401.15328
-
Fortifying Ethical Boundaries in AI: Advanced Strategies for Enhancing Security in Large Language Models 27 Jan 2024 · 0 repositories · arXiv:2402.01725
-
GeoDecoder: Empowering Multimodal Map Understanding 26 Jan 2024 · 0 repositories · arXiv:2401.15118
-
Scalable Qualitative Coding with LLMs: Chain-of-Thought Reasoning Matches Human Performance in Some Hermeneutic Tasks 26 Jan 2024 · 0 repositories · arXiv:2401.15170
-
A comparative study of zero-shot inference with large language models and supervised modeling in breast cancer pathology classification 25 Jan 2024 · 0 repositories · arXiv:2401.13887
-
(Chat)GPT v BERT: Dawn of Justice for Semantic Change Detection 25 Jan 2024 · 1 repository · arXiv:2401.14040
-
DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence 25 Jan 2024 · 1 repository · arXiv:2401.14196Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
Evaluating GPT-3.5's Awareness and Summarization Abilities for European Constitutional Texts with Shared Topics 25 Jan 2024 · 0 repositories · arXiv:2401.14524
-
Investigate-Consolidate-Exploit: A General Strategy for Inter-Task Agent Self-Evolution 25 Jan 2024 · 0 repositories · arXiv:2401.13996
-
LongHealth: A Question Answering Benchmark with Long Clinical Documents 25 Jan 2024 · 1 repository · arXiv:2401.14490Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
TrICy: Trigger-guided Data-to-text Generation with Intent aware Attention-Copy 25 Jan 2024 · 0 repositories · arXiv:2402.01714
-
Unmasking and Quantifying Racial Bias of Large Language Models in Medical Report Generation 25 Jan 2024 · 0 repositories · arXiv:2401.13867
-
ZS4C: Zero-Shot Synthesis of Compilable Code for Incomplete Code Snippets using LLMs 25 Jan 2024 · 0 repositories · arXiv:2401.14279
-
A Unified Approach to Emotion Detection and Task-Oriented Dialogue Modeling 24 Jan 2024 · 1 repository · arXiv:2401.13789
-
Automated Root Causing of Cloud Incidents using In-Context Learning with GPT-4 24 Jan 2024 · 0 repositories · arXiv:2401.13810
-
Can GPT-3.5 Generate and Code Discharge Summaries? 24 Jan 2024 · 1 repository · arXiv:2401.13512Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Discovering Mathematical Formulas from Data via GPT-guided Monte Carlo Tree Search 24 Jan 2024 · 0 repositories · arXiv:2401.14424
-
Evaluation of General Large Language Models in Contextually Assessing Semantic Concepts Extracted from Adult Critical Care Electronic Health Record Notes 24 Jan 2024 · 0 repositories · arXiv:2401.13588
-
Graph Guided Question Answer Generation for Procedural Question-Answering 24 Jan 2024 · 0 repositories · arXiv:2401.13594
-
How Good is ChatGPT at Face Biometrics? A First Look into Recognition, Soft Biometrics, and Explainability 24 Jan 2024 · 1 repository · arXiv:2401.13641
-
SEER: Facilitating Structured Reasoning and Explanation via Reinforcement Learning 24 Jan 2024 · 1 repository · arXiv:2401.13246Syntology official (archive's flag): 12 ran · 13 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 5 where Syntology's instrument failed) · 3 unverified (of 16 harvested samples) · 2 pointer-only (licence)
-
Segment Any Cell: A SAM-based Auto-prompting Fine-tuning Framework for Nuclei Segmentation 24 Jan 2024 · 0 repositories · arXiv:2401.13220
-
KAM-CoT: Knowledge Augmented Multimodal Chain-of-Thoughts Reasoning 23 Jan 2024 · 0 repositories · arXiv:2401.12863
-
TroVE: Inducing Verifiable and Efficient Toolboxes for Solving Programmatic Tasks 23 Jan 2024 · 1 repository · arXiv:2401.12869Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 16 with no instrument failure: 0 honoured, 0 violated, 16 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 19 harvested samples) · 19 pointer-only (licence)
-
Enhancing In-context Learning via Linear Probe Calibration 22 Jan 2024 · 1 repository · arXiv:2401.12406Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
CheX-GPT: Harnessing Large Language Models for Enhanced Chest X-ray Report Labeling 21 Jan 2024 · 2 repositories · arXiv:2401.11505
-
Enhancing Recommendation Diversity by Re-ranking with Large Language Models 21 Jan 2024 · 0 repositories · arXiv:2401.11506
-
BadChain: Backdoor Chain-of-Thought Prompting for Large Language Models 20 Jan 2024 · 1 repository · arXiv:2401.12242Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 15 harvested samples)
-
Enhancing Large Language Models for Clinical Decision Support by Incorporating Clinical Practice Guidelines 20 Jan 2024 · 0 repositories · arXiv:2401.11120
-
Evaluating and Enhancing Large Language Models Performance in Domain-specific Medicine: Osteoarthritis Management with DocOA 20 Jan 2024 · 0 repositories · arXiv:2401.12998
-
FinLLMs: A Framework for Financial Reasoning Dataset Generation with Large Language Models 19 Jan 2024 · 0 repositories · arXiv:2401.10744
-
Mining experimental data from Materials Science literature with Large Language Models: an evaluation study 19 Jan 2024 · 1 repository · arXiv:2401.11052
-
Reinforcement learning for question answering in programming domain using public community scoring as a human feedback 19 Jan 2024 · 0 repositories · arXiv:2401.10882
-
ChatQA: Surpassing GPT-4 on Conversational QA and RAG 18 Jan 2024 · 0 repositories · arXiv:2401.10225
-
Code Prompting Elicits Conditional Reasoning Abilities in Text+Code LLMs 18 Jan 2024 · 1 repository · arXiv:2401.10065Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Gender Bias in Machine Translation and The Era of Large Language Models 18 Jan 2024 · 0 repositories · arXiv:2401.10016
-
Image Translation as Diffusion Visual Programmers 18 Jan 2024 · 0 repositories · arXiv:2401.09742
-
Leveraging Biases in Large Language Models: "bias-kNN'' for Effective Few-Shot Learning 18 Jan 2024 · 0 repositories · arXiv:2401.09783
-
When Neural Code Completion Models Size up the Situation: Attaining Cheaper and Faster Completion through Dynamic Model Inference 18 Jan 2024 · 1 repository · arXiv:2401.09964
-
Improving Classification Performance With Human Feedback: Label a few, we label the rest 17 Jan 2024 · 0 repositories · arXiv:2401.09555
-
Learning from Implicit User Feedback, Emotions and Demographic Information in Task-Oriented and Document-Grounded Dialogues 17 Jan 2024 · 1 repository · arXiv:2401.09248
-
MADA: Meta-Adaptive Optimizers through hyper-gradient Descent 17 Jan 2024 · 0 repositories · arXiv:2401.08893
-
Application of LLM Agents in Recruitment: A Novel Framework for Resume Screening 16 Jan 2024 · 0 repositories · arXiv:2401.08315
-
Enhancing Robustness of LLM-Synthetic Text Detectors for Academic Writing: A Comprehensive Analysis 16 Jan 2024 · 0 repositories · arXiv:2401.08046