Methods › General › Learning Rate Schedules › Cosine Annealing › Papers, page 11
Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,965 · with a code link: 1,734 · where Syntology ran a sample: 627 (513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (627 of 3,965 tagged: 513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument)
Page 11 of 40: papers 1,001 to 1,100 of 3,965, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
PlagBench: Exploring the Duality of Large Language Models in Plagiarism Generation and Detection 24 Jun 2024 · 0 repositories · arXiv:2406.16288
-
The GPT-WritingPrompts Dataset: A Comparative Analysis of Character Portrayal in Short Stories 24 Jun 2024 · 1 repository · arXiv:2406.16767
-
Towards Better Graph-based Cross-document Relation Extraction via Non-bridge Entity Enhancement and Prediction Debiasing 24 Jun 2024 · 1 repository · arXiv:2406.16529
-
Evaluating the Effectiveness of the Foundational Models for Q&A Classification in Mental Health care 23 Jun 2024 · 0 repositories · arXiv:2406.15966
-
GraphEval2000: Benchmarking and Improving Large Language Models on Graph Datasets 23 Jun 2024 · 0 repositories · arXiv:2406.16176
-
Can LLMs Generate Visualizations with Dataless Prompts? 22 Jun 2024 · 0 repositories · arXiv:2406.17805
-
SS-GEN: A Social Story Generation Framework with Large Language Models 22 Jun 2024 · 2 repositories · arXiv:2406.15695
-
A GPT-based Code Review System for Programming Language Learning 21 Jun 2024 · 0 repositories · arXiv:2407.04722
-
Anime Popularity Prediction Before Huge Investments: a Multimodal Approach Using Deep Learning 21 Jun 2024 · 0 repositories · arXiv:2406.16961
-
ChatGPT as Research Scientist: Probing GPT's Capabilities as a Research Librarian, Research Ethicist, Data Generator and Data Predictor 20 Jun 2024 · 0 repositories · arXiv:2406.14765
-
CryptoGPT: a 7B model rivaling GPT-4 in the task of analyzing and classifying real-time financial news 20 Jun 2024 · 0 repositories · arXiv:2406.14039
-
Evaluating Implicit Bias in Large Language Models by Attacking From a Psychometric Perspective 20 Jun 2024 · 1 repository · arXiv:2406.14023
-
Generative AI for Enhancing Active Learning in Education: A Comparative Study of GPT-3.5 and GPT-4 in Crafting Customized Test Questions 20 Jun 2024 · 0 repositories · arXiv:2406.13903
-
How to Compute the Probability of a Word 20 Jun 2024 · 2 repositories · arXiv:2406.14561Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Identifying User Goals from UI Trajectories 20 Jun 2024 · 0 repositories · arXiv:2406.14314
-
Learning to Plan for Retrieval-Augmented Large Language Models from Knowledge Graphs 20 Jun 2024 · 1 repository · arXiv:2406.14282Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
LLaSA: A Multimodal LLM for Human Activity Analysis Through Wearable and Smartphone Sensors 20 Jun 2024 · 1 repository · arXiv:2406.14498Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Persuasiveness of Generated Free-Text Rationales in Subjective Decisions: A Case Study on Pairwise Argument Ranking 20 Jun 2024 · 1 repository · arXiv:2406.13905
-
Prism: A Framework for Decoupling and Assessing the Capabilities of VLMs 20 Jun 2024 · 1 repository · arXiv:2406.14544Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples)
-
SciDMT: A Large-Scale Corpus for Detecting Scientific Mentions 20 Jun 2024 · 0 repositories · arXiv:2406.14756
-
On AI-Inspired UI-Design 19 Jun 2024 · 1 repository · arXiv:2406.13631
-
Open Generative Large Language Models for Galician 19 Jun 2024 · 0 repositories · arXiv:2406.13893
-
Part-aware Unified Representation of Language and Skeleton for Zero-shot Action Recognition 19 Jun 2024 · 1 repository · arXiv:2406.13327Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Generating Educational Materials with Different Levels of Readability using LLMs 18 Jun 2024 · 0 repositories · arXiv:2406.12787
-
IPEval: A Bilingual Intellectual Property Agency Consultation Evaluation Benchmark for Large Language Models 18 Jun 2024 · 1 repository · arXiv:2406.12386
-
Towards a Client-Centered Assessment of LLM Therapists by Client Simulation 18 Jun 2024 · 1 repository · arXiv:2406.12266Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
UrbanLLM: Autonomous Urban Activity Planning and Management with Large Language Models 18 Jun 2024 · 0 repositories · arXiv:2406.12360
-
Vernacular? I Barely Know Her: Challenges with Style Control and Stereotyping 18 Jun 2024 · 0 repositories · arXiv:2406.12679
-
What Makes Two Language Models Think Alike? 18 Jun 2024 · 0 repositories · arXiv:2406.12620
-
"You Gotta be a Doctor, Lin": An Investigation of Name-Based Bias of Large Language Models in Employment Recommendations 18 Jun 2024 · 0 repositories · arXiv:2406.12232
-
Breaking Boundaries: Investigating the Effects of Model Editing on Cross-linguistic Performance 17 Jun 2024 · 0 repositories · arXiv:2406.11139
-
Building another Spanish dictionary, this time with GPT-4 17 Jun 2024 · 1 repository · arXiv:2406.11218
-
Cultural Conditioning or Placebo? On the Effectiveness of Socio-Demographic Prompting 17 Jun 2024 · 0 repositories · arXiv:2406.11661
-
SeRTS: Self-Rewarding Tree Search for Biomedical Retrieval-Augmented Generation 17 Jun 2024 · 0 repositories · arXiv:2406.11258
-
Enhancing Text Classification through LLM-Driven Active Learning and Human Annotation 17 Jun 2024 · 1 repository · arXiv:2406.12114
-
Estimating the Increase in Emissions caused by AI-augmented Search 17 Jun 2024 · 0 repositories · arXiv:2407.16894
-
Are Small Language Models Ready to Compete with Large Language Models for Practical Applications? 17 Jun 2024 · 1 repository · arXiv:2406.11402
-
Exploring Safety-Utility Trade-Offs in Personalized Language Models 17 Jun 2024 · 0 repositories · arXiv:2406.11107
-
GPT-Powered Elicitation Interview Script Generator for Requirements Engineering Training 17 Jun 2024 · 0 repositories · arXiv:2406.11439
-
Improving Multi-Agent Debate with Sparse Communication Topology 17 Jun 2024 · 0 repositories · arXiv:2406.11776
-
Investigating Annotator Bias in Large Language Models for Hate Speech Detection 17 Jun 2024 · 3 repositories · arXiv:2406.11109
-
JobFair: A Framework for Benchmarking Gender Hiring Bias in Large Language Models 17 Jun 2024 · 0 repositories · arXiv:2406.15484
-
Promises, Outlooks and Challenges of Diffusion Language Modeling 17 Jun 2024 · 0 repositories · arXiv:2406.11473
-
Scaling the Codebook Size of VQGAN to 100,000 with a Utilization Rate of 99% 17 Jun 2024 · 1 repository · arXiv:2406.11837Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Self and Cross-Model Distillation for LLMs: Effective Methods for Refusal Pattern Alignment 17 Jun 2024 · 0 repositories · arXiv:2406.11285
-
WellDunn: On the Robustness and Explainability of Language Models and Large Language Models in Identifying Wellness Dimensions 17 Jun 2024 · 1 repository · arXiv:2406.12058Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Enhancing Supermarket Robot Interaction: A Multi-Level LLM Conversational Interface for Handling Diverse Customer Intents 16 Jun 2024 · 0 repositories · arXiv:2406.11047
-
Exposing the Achilles' Heel: Evaluating LLMs Ability to Handle Mistakes in Mathematical Reasoning 16 Jun 2024 · 0 repositories · arXiv:2406.10834
-
Generating Tables from the Parametric Knowledge of Language Models 16 Jun 2024 · 1 repository · arXiv:2406.10922
-
Grading Massive Open Online Courses Using Large Language Models 16 Jun 2024 · 0 repositories · arXiv:2406.11102
-
KGPA: Robustness Evaluation for Large Language Models via Cross-Domain Knowledge Graphs 16 Jun 2024 · 1 repository · arXiv:2406.10802
-
Large Language Models for Automatic Milestone Detection in Group Discussions 16 Jun 2024 · 0 repositories · arXiv:2406.10842
-
ShareLoRA: Parameter Efficient and Robust Large Language Model Fine-tuning via Shared Low-Rank Adaptation 16 Jun 2024 · 1 repository · arXiv:2406.10785Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
ViD-GPT: Introducing GPT-style Autoregressive Generation in Video Diffusion Models 16 Jun 2024 · 1 repository · arXiv:2406.10981
-
A Comprehensive Survey of Foundation Models in Medicine 15 Jun 2024 · 0 repositories · arXiv:2406.10729
-
Beyond Raw Videos: Understanding Edited Videos with Large Multimodal Model 15 Jun 2024 · 1 repository · arXiv:2406.10484
-
MINT: a Multi-modal Image and Narrative Text Dubbing Dataset for Foley Audio Content Planning and Generation 15 Jun 2024 · 1 repository · arXiv:2406.10591
-
Exploring the Correlation between Human and Machine Evaluation of Simultaneous Speech Translation 14 Jun 2024 · 0 repositories · arXiv:2406.10091
-
Towards Efficient Pareto Set Approximation via Mixture of Experts Based Model Fusion 14 Jun 2024 · 1 repository · arXiv:2406.09770
-
A More Practical Approach to Machine Unlearning 13 Jun 2024 · 0 repositories · arXiv:2406.09391
-
Linguistic Bias in ChatGPT: Language Models Reinforce Dialect Discrimination 13 Jun 2024 · 0 repositories · arXiv:2406.08818
-
Optimizing Large Model Training through Overlapped Activation Recomputation 13 Jun 2024 · 0 repositories · arXiv:2406.08756
-
Talking Heads: Understanding Inter-layer Communication in Transformer Language Models 13 Jun 2024 · 0 repositories · arXiv:2406.09519
-
FaithFill: Faithful Inpainting for Object Completion Using a Single Reference Image 12 Jun 2024 · 0 repositories · arXiv:2406.07865
-
Fine-Tuned 'Small' LLMs (Still) Significantly Outperform Zero-Shot Generative AI Models in Text Classification 12 Jun 2024 · 1 repository · arXiv:2406.08660Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
How well it works: Benchmarking performance of GPT models on medical natural language processing tasks 12 Jun 2024 · 0 repositories
-
Making Task-Oriented Dialogue Datasets More Natural by Synthetically Generating Indirect User Requests 12 Jun 2024 · 0 repositories · arXiv:2406.07794
-
Tailoring Generative AI Chatbots for Multiethnic Communities in Disaster Preparedness Communication: Extending the CASA Paradigm 12 Jun 2024 · 1 repository · arXiv:2406.08411
-
Beyond Words: On Large Language Models Actionability in Mission-Critical Risk Analysis 11 Jun 2024 · 0 repositories · arXiv:2406.10273
-
Bilingual Sexism Classification: Fine-Tuned XLM-RoBERTa and GPT-3.5 Few-Shot Learning 11 Jun 2024 · 0 repositories · arXiv:2406.07287
-
Flextron: Many-in-One Flexible Large Language Model 11 Jun 2024 · 0 repositories · arXiv:2406.10260
-
Multi-objective Reinforcement learning from AI Feedback 11 Jun 2024 · 1 repository · arXiv:2406.07295
-
Unused information in token probability distribution of generative LLM: improving LLM reading comprehension through calculation of expected values 11 Jun 2024 · 1 repository · arXiv:2406.10267
-
AGB-DE: A Corpus for the Automated Legal Assessment of Clauses in German Consumer Contracts 10 Jun 2024 · 1 repository · arXiv:2406.06809
-
Compute Better Spent: Replacing Dense Layers with Structured Matrices 10 Jun 2024 · 1 repository · arXiv:2406.06248Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
In-Context Learning and Fine-Tuning GPT for Argument Mining 10 Jun 2024 · 1 repository · arXiv:2406.06699
-
LLM-dCache: Improving Tool-Augmented LLMs with GPT-Driven Localized Data Caching 10 Jun 2024 · 0 repositories · arXiv:2406.06799
-
SecureNet: A Comparative Study of DeBERTa and Large Language Models for Phishing Detection 10 Jun 2024 · 0 repositories · arXiv:2406.06663
-
DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation 9 Jun 2024 · 2 repositories · arXiv:2406.05654Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Hidden Holes: topological aspects of language models 9 Jun 2024 · 0 repositories · arXiv:2406.05798
-
MedREQAL: Examining Medical Knowledge Recall of Large Language Models via Question Answering 9 Jun 2024 · 0 repositories · arXiv:2406.05845
-
Text2VP: Generative AI for Visual Programming and Parametric Modeling 9 Jun 2024 · 0 repositories · arXiv:2407.07732
-
Critical Phase Transition in Large Language Models 8 Jun 2024 · 0 repositories · arXiv:2406.05335
-
Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts 8 Jun 2024 · 0 repositories · arXiv:2406.05569
-
MaTableGPT: GPT-based Table Data Extractor from Materials Science Literature 8 Jun 2024 · 0 repositories · arXiv:2406.05431
-
SelfDefend: LLMs Can Defend Themselves against Jailbreaking in a Practical Manner 8 Jun 2024 · 0 repositories · arXiv:2406.05498
-
BAMO at SemEval-2024 Task 9: BRAINTEASER: A Novel Task Defying Common Sense 7 Jun 2024 · 1 repository · arXiv:2406.04947
-
BERTs are Generative In-Context Learners 7 Jun 2024 · 1 repository · arXiv:2406.04823Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 1 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 13 unverified (of 26 harvested samples)
-
GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents 7 Jun 2024 · 2 repositories · arXiv:2406.06613Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Large Generative Graph Models 7 Jun 2024 · 0 repositories · arXiv:2406.05109
-
Low-Resource Cross-Lingual Summarization through Few-Shot Learning with Large Language Models 7 Jun 2024 · 0 repositories · arXiv:2406.04630
-
On Subjective Uncertainty Quantification and Calibration in Natural Language Generation 7 Jun 2024 · 1 repository · arXiv:2406.05213Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples)
-
VTrans: Accelerating Transformer Compression with Variational Information Bottleneck based Pruning 7 Jun 2024 · 0 repositories · arXiv:2406.05276
-
A Survey on Medical Large Language Models: Technology, Application, Trustworthiness, and Future Directions 6 Jun 2024 · 0 repositories · arXiv:2406.03712
-
Do Language Models Understand Morality? Towards a Robust Detection of Moral Content 6 Jun 2024 · 1 repository · arXiv:2406.04143
-
HORAE: A Domain-Agnostic Language for Automated Service Regulation 6 Jun 2024 · 1 repository · arXiv:2406.06600
-
LLMEmbed: Rethinking Lightweight LLM's Genuine Function in Text Classification 6 Jun 2024 · 1 repository · arXiv:2406.03725
-
Simplified and Generalized Masked Diffusion for Discrete Data 6 Jun 2024 · 1 repository · arXiv:2406.04329Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 15 unverified (of 21 harvested samples)
-
Tox-BART: Leveraging Toxicity Attributes for Explanation Generation of Implicit Hate Speech 6 Jun 2024 · 1 repository · arXiv:2406.03953
-
Your Absorbing Discrete Diffusion Secretly Models the Conditional Distributions of Clean Data 6 Jun 2024 · 2 repositories · arXiv:2406.03736Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples) · 13 pointer-only (licence)