Methods › General › Learning Rate Schedules › Cosine Annealing › Papers, page 10
Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,965 · with a code link: 1,734 · where Syntology ran a sample: 627 (513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (627 of 3,965 tagged: 513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument)
Page 10 of 40: papers 901 to 1,000 of 3,965, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Evaluating Large Language Models for Anxiety and Depression Classification using Counseling and Psychotherapy Transcripts 18 Jul 2024 · 1 repository · arXiv:2407.13228
-
How Reliable are LLMs as Knowledge Bases? Re-thinking Facutality and Consistency 18 Jul 2024 · 0 repositories · arXiv:2407.13578
-
Learning-From-Mistakes Prompting for Indigenous Language Translation 18 Jul 2024 · 0 repositories · arXiv:2407.13343
-
PRAGyan -- Connecting the Dots in Tweets 18 Jul 2024 · 0 repositories · arXiv:2407.13909
-
Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction 18 Jul 2024 · 1 repository · arXiv:2407.13943Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Beyond Binary: Multiclass Paraphasia Detection with Generative Pretrained Transformers and End-to-End Models 16 Jul 2024 · 0 repositories · arXiv:2407.11345
-
ChatBCG: Can AI Read Your Slide Deck? 16 Jul 2024 · 0 repositories · arXiv:2407.12875
-
Does Refusal Training in LLMs Generalize to the Past Tense? 16 Jul 2024 · 1 repository · arXiv:2407.11969Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
GPT Assisted Annotation of Rhetorical and Linguistic Features for Interpretable Propaganda Technique Detection in News Text 16 Jul 2024 · 0 repositories · arXiv:2407.11827
-
LaMI-DETR: Open-Vocabulary Detection with Language Model Instruction 16 Jul 2024 · 1 repository · arXiv:2407.11335Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Large Language Models as Misleading Assistants in Conversation 16 Jul 2024 · 0 repositories · arXiv:2407.11789
-
Large Visual-Language Models Are Also Good Classifiers: A Study of In-Context Multimodal Fake News Detection 16 Jul 2024 · 0 repositories · arXiv:2407.12879
-
Representation Bias in Political Sample Simulations with Large Language Models 16 Jul 2024 · 0 repositories · arXiv:2407.11409
-
ReFeR: Improving Evaluation and Reasoning through Hierarchy of Models 16 Jul 2024 · 0 repositories · arXiv:2407.12877
-
Trust No Bot: Discovering Personal Disclosures in Human-LLM Conversations in the Wild 16 Jul 2024 · 1 repository · arXiv:2407.11438
-
CodeV: Empowering LLMs with HDL Generation through Multi-Level Summarization 15 Jul 2024 · 0 repositories · arXiv:2407.10424
-
Leveraging LLM-Respondents for Item Evaluation: a Psychometric Analysis 15 Jul 2024 · 0 repositories · arXiv:2407.10899
-
Making New Connections: LLMs as Puzzle Generators for The New York Times' Connections Word Game 15 Jul 2024 · 0 repositories · arXiv:2407.11240
-
Mechanistic interpretability of large language models with applications to the financial services industry 15 Jul 2024 · 0 repositories · arXiv:2407.11215
-
MetaLLM: A High-performant and Cost-efficient Dynamic Framework for Wrapping LLMs 15 Jul 2024 · 1 repository · arXiv:2407.10834Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Think-on-Graph 2.0: Deep and Faithful Large Language Model Reasoning with Knowledge-guided Retrieval Augmented Generation 15 Jul 2024 · 1 repository · arXiv:2407.10805Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 13 with no instrument failure: 1 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 16 harvested samples)
-
Curriculum Learning for Small Code Language Models 14 Jul 2024 · 0 repositories · arXiv:2407.10194
-
Document-level Clinical Entity and Relation Extraction via Knowledge Base-Guided Generation 13 Jul 2024 · 0 repositories · arXiv:2407.10021
-
Generating In-store Customer Journeys from Scratch with GPT Architectures 13 Jul 2024 · 0 repositories · arXiv:2407.11081
-
ASTPrompter: Weakly Supervised Automated Language Model Red-Teaming to Identify Low-Perplexity Toxic Prompts 12 Jul 2024 · 1 repository · arXiv:2407.09447
-
Self-Evolving GPT: A Lifelong Autonomous Experiential Learner 12 Jul 2024 · 0 repositories · arXiv:2407.08937
-
Show, Don't Tell: Evaluating Large Language Models Beyond Textual Understanding with ChildPlay 12 Jul 2024 · 1 repository · arXiv:2407.11068
-
The Two Sides of the Coin: Hallucination Generation and Detection with LLMs as Evaluators for LLMs 12 Jul 2024 · 0 repositories · arXiv:2407.09152
-
Toward Automatic Group Membership Annotation for Group Fairness Evaluation 12 Jul 2024 · 0 repositories · arXiv:2407.08926
-
GPT-4 is judged more human than humans in displaced and inverted Turing tests 11 Jul 2024 · 0 repositories · arXiv:2407.08853
-
HDT: Hierarchical Document Transformer 11 Jul 2024 · 0 repositories · arXiv:2407.08330
-
LLMs' morphological analyses of complex FST-generated Finnish words 11 Jul 2024 · 1 repository · arXiv:2407.08269
-
MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine 11 Jul 2024 · 3 repositories · arXiv:2407.08739
-
On the (In)Security of LLM App Stores 11 Jul 2024 · 0 repositories · arXiv:2407.08422
-
Real-Time Anomaly Detection and Reactive Planning with Large Language Models 11 Jul 2024 · 0 repositories · arXiv:2407.08735
-
Vox Populi, Vox AI? Using Language Models to Estimate German Public Opinion 11 Jul 2024 · 1 repository · arXiv:2407.08563
-
FsPONER: Few-shot Prompt Optimization for Named Entity Recognition in Domain-specific Scenarios 10 Jul 2024 · 1 repository · arXiv:2407.08035
-
KpopMT: Translation Dataset with Terminology for Kpop Fandom 10 Jul 2024 · 1 repository · arXiv:2407.07413
-
A Guide To Effectively Leveraging LLMs for Low-Resource Text Summarization: Data Augmentation and Semi-supervised Approaches 10 Jul 2024 · 0 repositories · arXiv:2407.07341
-
Multilingual Blending: LLM Safety Alignment Evaluation with Language Mixture 10 Jul 2024 · 0 repositories · arXiv:2407.07342
-
ROSA: Random Subspace Adaptation for Efficient Fine-Tuning 10 Jul 2024 · 1 repository · arXiv:2407.07802
-
AI AI Bias: Large Language Models Favor Their Own Generated Content 9 Jul 2024 · 1 repository · arXiv:2407.12856
-
ChatGPT Doesn't Trust Chargers Fans: Guardrail Sensitivity in Context 9 Jul 2024 · 1 repository · arXiv:2407.06866Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Identification of emotions on Twitter during the 2022 electoral process in Colombia 9 Jul 2024 · 0 repositories · arXiv:2407.07258
-
Measuring Sustainability Intention of ESG Fund Disclosure using Few-Shot Learning 9 Jul 2024 · 0 repositories · arXiv:2407.06893
-
Prompting Techniques for Secure Code Generation: A Systematic Investigation 9 Jul 2024 · 0 repositories · arXiv:2407.07064
-
Raply: A profanity-mitigated rap generator 9 Jul 2024 · 0 repositories · arXiv:2407.06941
-
Solving General Natural-Language-Description Optimization Problems with Large Language Models 9 Jul 2024 · 0 repositories · arXiv:2407.07924
-
Using Large Language Models for Generating Smart Contracts for Health Insurance from Textual Policies 9 Jul 2024 · 0 repositories · arXiv:2407.07019
-
Using Pretrained Large Language Model with Prompt Engineering to Answer Biomedical Questions 9 Jul 2024 · 0 repositories · arXiv:2407.06779
-
InverseCoder: Self-improving Instruction-Tuned Code LLMs with Inverse-Instruct 8 Jul 2024 · 1 repository · arXiv:2407.05700Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Large Language Models Understand Layout 8 Jul 2024 · 1 repository · arXiv:2407.05750
-
Potential of Multimodal Large Language Models for Data Mining of Medical Images and Free-text Reports 8 Jul 2024 · 0 repositories · arXiv:2407.05758
-
Surprising gender biases in GPT 8 Jul 2024 · 0 repositories · arXiv:2407.06003
-
SHINE: Saliency-aware HIerarchical NEgative Ranking for Compositional Temporal Grounding 6 Jul 2024 · 1 repository · arXiv:2407.05118Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 1 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Are Large Language Models Strategic Decision Makers? A Study of Performance and Bias in Two-Player Non-Zero-Sum Games 5 Jul 2024 · 0 repositories · arXiv:2407.04467
-
GPT vs RETRO: Exploring the Intersection of Retrieval and Parameter-Efficient Fine-Tuning 5 Jul 2024 · 0 repositories · arXiv:2407.04528
-
From Data to Commonsense Reasoning: The Use of Large Language Models for Explainable AI 4 Jul 2024 · 0 repositories · arXiv:2407.03778
-
NutriBench: A Dataset for Evaluating Large Language Models on Nutrition Estimation from Meal Descriptions 4 Jul 2024 · 0 repositories · arXiv:2407.12843
-
Controllable Conversations: Planning-Based Dialogue Agent with Large Language Models 4 Jul 2024 · 1 repository · arXiv:2407.03884
-
Question-Analysis Prompting Improves LLM Performance in Reasoning Tasks 4 Jul 2024 · 0 repositories · arXiv:2407.03624
-
Slice-100K: A Multimodal Dataset for Extrusion-based 3D Printing 4 Jul 2024 · 0 repositories · arXiv:2407.04180
-
Towards Automating Text Annotation: A Case Study on Semantic Proximity Annotation using GPT-4 4 Jul 2024 · 0 repositories · arXiv:2407.04130
-
AgentInstruct: Toward Generative Teaching with Agentic Flows 3 Jul 2024 · 0 repositories · arXiv:2407.03502
-
Gradient descent with generalized Newton's method 3 Jul 2024 · 1 repository · arXiv:2407.02772Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Improving LLM Abilities in Idiomatic Translation 3 Jul 2024 · 0 repositories · arXiv:2407.03518
-
ObfuscaTune: Obfuscated Offsite Fine-tuning and Inference of Proprietary LLMs on Private Datasets 3 Jul 2024 · 0 repositories · arXiv:2407.02960
-
OSPC: Artificial VLM Features for Hateful Meme Detection 3 Jul 2024 · 0 repositories · arXiv:2407.12836
-
Regurgitative Training: The Value of Real Data in Training Large Language Models 3 Jul 2024 · 0 repositories · arXiv:2407.12835
-
Assessing the Code Clone Detection Capability of Large Language Models 2 Jul 2024 · 0 repositories · arXiv:2407.02402
-
Beyond Numeric Awards: In-Context Dueling Bandits with LLM Agents 2 Jul 2024 · 0 repositories · arXiv:2407.01887
-
GPTCast: a weather language model for precipitation nowcasting 2 Jul 2024 · 1 repository · arXiv:2407.02089
-
GRASP: A Grid-Based Benchmark for Evaluating Commonsense Spatial Reasoning 2 Jul 2024 · 0 repositories · arXiv:2407.01892
-
Integrate the Essence and Eliminate the Dross: Fine-Grained Self-Consistency for Free-Form Language Generation 2 Jul 2024 · 1 repository · arXiv:2407.02056Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Language Model Alignment in Multilingual Trolley Problems 2 Jul 2024 · 2 repositories · arXiv:2407.02273Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
SeqAR: Jailbreak LLMs with Sequential Auto-Generated Characters 2 Jul 2024 · 1 repository · arXiv:2407.01902Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Increasing Model Capacity for Free: A Simple Strategy for Parameter Efficient Fine-tuning 1 Jul 2024 · 1 repository · arXiv:2407.01320Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 7 pointer-only (licence)
-
Predicting DC-Link Capacitor Current Ripple in AC-DC Rectifier Circuits Using Fine-Tuned Large Language Models 1 Jul 2024 · 0 repositories · arXiv:2407.01724
-
Parm: Efficient Training of Large Sparsely-Activated Models with Dedicated Schedules 30 Jun 2024 · 1 repository · arXiv:2407.00599
-
Applying RLAIF for Code Generation with API-usage in Lightweight LLMs 28 Jun 2024 · 0 repositories · arXiv:2406.20060
-
FRED: Flexible REduction-Distribution Interconnect and Communication Implementation for Wafer-Scale Distributed Training of DNN Models 28 Jun 2024 · 0 repositories · arXiv:2406.19580
-
Machine Learning Predictors for Min-Entropy Estimation 28 Jun 2024 · 1 repository · arXiv:2406.19983
-
ScaleBiO: Scalable Bilevel Optimization for LLM Data Reweighting 28 Jun 2024 · 0 repositories · arXiv:2406.19976
-
ShortcutsBench: A Large-Scale Real-world Benchmark for API-based Agents 28 Jun 2024 · 1 repository · arXiv:2407.00132Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Fibottention: Inceptive Visual Representation Learning with Diverse Attention Across Heads 27 Jun 2024 · 1 repository · arXiv:2406.19391
-
Fine-tuned network relies on generic representation to solve unseen cognitive task 27 Jun 2024 · 0 repositories · arXiv:2406.18926
-
From Artificial Needles to Real Haystacks: Improving Retrieval Capabilities in LLMs by Finetuning on Synthetic Data 27 Jun 2024 · 1 repository · arXiv:2406.19292
-
Granite-Function Calling Model: Introducing Function Calling Abilities via Multi-task Learning of Granular Tasks 27 Jun 2024 · 0 repositories · arXiv:2407.00121
-
The Model Arena for Cross-lingual Sentiment Analysis: A Comparative Study in the Era of Large Language Models 27 Jun 2024 · 0 repositories · arXiv:2406.19358
-
APIGen: Automated Pipeline for Generating Verifiable and Diverse Function-Calling Datasets 26 Jun 2024 · 0 repositories · arXiv:2406.18518
-
FactFinders at CheckThat! 2024: Refining Check-worthy Statement Detection with LLMs through Data Pruning 26 Jun 2024 · 1 repository · arXiv:2406.18297
-
Improving Entity Recognition Using Ensembles of Deep Learning and Fine-tuned Large Language Models: A Case Study on Adverse Event Extraction from Multiple Sources 26 Jun 2024 · 0 repositories · arXiv:2406.18049
-
MathOdyssey: Benchmarking Mathematical Problem-Solving Skills in Large Language Models Using Odyssey Math Data 26 Jun 2024 · 3 repositories · arXiv:2406.18321Syntology 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
SetBERT: Enhancing Retrieval Performance for Boolean Logic and Set Operation Queries 25 Jun 2024 · 0 repositories · arXiv:2406.17282
-
Interpreting Attention Layer Outputs with Sparse Autoencoders 25 Jun 2024 · 1 repository · arXiv:2406.17759Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Understanding Language Model Circuits through Knowledge Editing 25 Jun 2024 · 0 repositories · arXiv:2406.17241
-
DreamBench++: A Human-Aligned Benchmark for Personalized Image Generation 24 Jun 2024 · 1 repository · arXiv:2406.16855Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Evaluation of Instruction-Following Ability for Large Language Models on Story-Ending Generation 24 Jun 2024 · 0 repositories · arXiv:2406.16356
-
Finding Transformer Circuits with Edge Pruning 24 Jun 2024 · 1 repository · arXiv:2406.16778Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples)
-
modeLing: A Novel Dataset for Testing Linguistic Reasoning in Language Models 24 Jun 2024 · 0 repositories · arXiv:2406.17038