Methods › General › Learning Rate Schedules › Linear Warmup With Cosine Annealing › Papers, page 11
Linear Warmup With Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,797 · with a code link: 1,655 · where Syntology ran a sample: 602 (490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (602 of 3,797 tagged: 490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument)
Page 11 of 38: papers 1,001 to 1,100 of 3,797, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
On AI-Inspired UI-Design 19 Jun 2024 · 1 repository · arXiv:2406.13631
-
Open Generative Large Language Models for Galician 19 Jun 2024 · 0 repositories · arXiv:2406.13893
-
Part-aware Unified Representation of Language and Skeleton for Zero-shot Action Recognition 19 Jun 2024 · 1 repository · arXiv:2406.13327Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 1 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Generating Educational Materials with Different Levels of Readability using LLMs 18 Jun 2024 · 0 repositories · arXiv:2406.12787
-
IPEval: A Bilingual Intellectual Property Agency Consultation Evaluation Benchmark for Large Language Models 18 Jun 2024 · 1 repository · arXiv:2406.12386
-
Towards a Client-Centered Assessment of LLM Therapists by Client Simulation 18 Jun 2024 · 1 repository · arXiv:2406.12266Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
UrbanLLM: Autonomous Urban Activity Planning and Management with Large Language Models 18 Jun 2024 · 0 repositories · arXiv:2406.12360
-
Vernacular? I Barely Know Her: Challenges with Style Control and Stereotyping 18 Jun 2024 · 0 repositories · arXiv:2406.12679
-
What Makes Two Language Models Think Alike? 18 Jun 2024 · 0 repositories · arXiv:2406.12620
-
"You Gotta be a Doctor, Lin": An Investigation of Name-Based Bias of Large Language Models in Employment Recommendations 18 Jun 2024 · 0 repositories · arXiv:2406.12232
-
Breaking Boundaries: Investigating the Effects of Model Editing on Cross-linguistic Performance 17 Jun 2024 · 0 repositories · arXiv:2406.11139
-
Building another Spanish dictionary, this time with GPT-4 17 Jun 2024 · 1 repository · arXiv:2406.11218
-
Cultural Conditioning or Placebo? On the Effectiveness of Socio-Demographic Prompting 17 Jun 2024 · 0 repositories · arXiv:2406.11661
-
SeRTS: Self-Rewarding Tree Search for Biomedical Retrieval-Augmented Generation 17 Jun 2024 · 0 repositories · arXiv:2406.11258
-
Enhancing Text Classification through LLM-Driven Active Learning and Human Annotation 17 Jun 2024 · 1 repository · arXiv:2406.12114
-
Estimating the Increase in Emissions caused by AI-augmented Search 17 Jun 2024 · 0 repositories · arXiv:2407.16894
-
Are Small Language Models Ready to Compete with Large Language Models for Practical Applications? 17 Jun 2024 · 1 repository · arXiv:2406.11402
-
Exploring Safety-Utility Trade-Offs in Personalized Language Models 17 Jun 2024 · 0 repositories · arXiv:2406.11107
-
GPT-Powered Elicitation Interview Script Generator for Requirements Engineering Training 17 Jun 2024 · 0 repositories · arXiv:2406.11439
-
Improving Multi-Agent Debate with Sparse Communication Topology 17 Jun 2024 · 0 repositories · arXiv:2406.11776
-
Investigating Annotator Bias in Large Language Models for Hate Speech Detection 17 Jun 2024 · 3 repositories · arXiv:2406.11109
-
JobFair: A Framework for Benchmarking Gender Hiring Bias in Large Language Models 17 Jun 2024 · 0 repositories · arXiv:2406.15484
-
Promises, Outlooks and Challenges of Diffusion Language Modeling 17 Jun 2024 · 0 repositories · arXiv:2406.11473
-
Scaling the Codebook Size of VQGAN to 100,000 with a Utilization Rate of 99% 17 Jun 2024 · 1 repository · arXiv:2406.11837Syntology official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Self and Cross-Model Distillation for LLMs: Effective Methods for Refusal Pattern Alignment 17 Jun 2024 · 0 repositories · arXiv:2406.11285
-
WellDunn: On the Robustness and Explainability of Language Models and Large Language Models in Identifying Wellness Dimensions 17 Jun 2024 · 1 repository · arXiv:2406.12058Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Enhancing Supermarket Robot Interaction: A Multi-Level LLM Conversational Interface for Handling Diverse Customer Intents 16 Jun 2024 · 0 repositories · arXiv:2406.11047
-
Exposing the Achilles' Heel: Evaluating LLMs Ability to Handle Mistakes in Mathematical Reasoning 16 Jun 2024 · 0 repositories · arXiv:2406.10834
-
Generating Tables from the Parametric Knowledge of Language Models 16 Jun 2024 · 1 repository · arXiv:2406.10922
-
Grading Massive Open Online Courses Using Large Language Models 16 Jun 2024 · 0 repositories · arXiv:2406.11102
-
KGPA: Robustness Evaluation for Large Language Models via Cross-Domain Knowledge Graphs 16 Jun 2024 · 1 repository · arXiv:2406.10802
-
Large Language Models for Automatic Milestone Detection in Group Discussions 16 Jun 2024 · 0 repositories · arXiv:2406.10842
-
ShareLoRA: Parameter Efficient and Robust Large Language Model Fine-tuning via Shared Low-Rank Adaptation 16 Jun 2024 · 1 repository · arXiv:2406.10785Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
ViD-GPT: Introducing GPT-style Autoregressive Generation in Video Diffusion Models 16 Jun 2024 · 1 repository · arXiv:2406.10981
-
A Comprehensive Survey of Foundation Models in Medicine 15 Jun 2024 · 0 repositories · arXiv:2406.10729
-
Beyond Raw Videos: Understanding Edited Videos with Large Multimodal Model 15 Jun 2024 · 1 repository · arXiv:2406.10484
-
MINT: a Multi-modal Image and Narrative Text Dubbing Dataset for Foley Audio Content Planning and Generation 15 Jun 2024 · 1 repository · arXiv:2406.10591
-
Exploring the Correlation between Human and Machine Evaluation of Simultaneous Speech Translation 14 Jun 2024 · 0 repositories · arXiv:2406.10091
-
Towards Efficient Pareto Set Approximation via Mixture of Experts Based Model Fusion 14 Jun 2024 · 1 repository · arXiv:2406.09770
-
A More Practical Approach to Machine Unlearning 13 Jun 2024 · 0 repositories · arXiv:2406.09391
-
Linguistic Bias in ChatGPT: Language Models Reinforce Dialect Discrimination 13 Jun 2024 · 0 repositories · arXiv:2406.08818
-
Optimizing Large Model Training through Overlapped Activation Recomputation 13 Jun 2024 · 0 repositories · arXiv:2406.08756
-
Talking Heads: Understanding Inter-layer Communication in Transformer Language Models 13 Jun 2024 · 0 repositories · arXiv:2406.09519
-
FaithFill: Faithful Inpainting for Object Completion Using a Single Reference Image 12 Jun 2024 · 0 repositories · arXiv:2406.07865
-
Fine-Tuned 'Small' LLMs (Still) Significantly Outperform Zero-Shot Generative AI Models in Text Classification 12 Jun 2024 · 1 repository · arXiv:2406.08660Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
How well it works: Benchmarking performance of GPT models on medical natural language processing tasks 12 Jun 2024 · 0 repositories
-
Making Task-Oriented Dialogue Datasets More Natural by Synthetically Generating Indirect User Requests 12 Jun 2024 · 0 repositories · arXiv:2406.07794
-
Tailoring Generative AI Chatbots for Multiethnic Communities in Disaster Preparedness Communication: Extending the CASA Paradigm 12 Jun 2024 · 1 repository · arXiv:2406.08411
-
Beyond Words: On Large Language Models Actionability in Mission-Critical Risk Analysis 11 Jun 2024 · 0 repositories · arXiv:2406.10273
-
Bilingual Sexism Classification: Fine-Tuned XLM-RoBERTa and GPT-3.5 Few-Shot Learning 11 Jun 2024 · 0 repositories · arXiv:2406.07287
-
Flextron: Many-in-One Flexible Large Language Model 11 Jun 2024 · 0 repositories · arXiv:2406.10260
-
Multi-objective Reinforcement learning from AI Feedback 11 Jun 2024 · 1 repository · arXiv:2406.07295
-
Unused information in token probability distribution of generative LLM: improving LLM reading comprehension through calculation of expected values 11 Jun 2024 · 1 repository · arXiv:2406.10267
-
AGB-DE: A Corpus for the Automated Legal Assessment of Clauses in German Consumer Contracts 10 Jun 2024 · 1 repository · arXiv:2406.06809
-
Compute Better Spent: Replacing Dense Layers with Structured Matrices 10 Jun 2024 · 1 repository · arXiv:2406.06248Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
In-Context Learning and Fine-Tuning GPT for Argument Mining 10 Jun 2024 · 1 repository · arXiv:2406.06699
-
LLM-dCache: Improving Tool-Augmented LLMs with GPT-Driven Localized Data Caching 10 Jun 2024 · 0 repositories · arXiv:2406.06799
-
SecureNet: A Comparative Study of DeBERTa and Large Language Models for Phishing Detection 10 Jun 2024 · 0 repositories · arXiv:2406.06663
-
DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation 9 Jun 2024 · 2 repositories · arXiv:2406.05654Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Hidden Holes: topological aspects of language models 9 Jun 2024 · 0 repositories · arXiv:2406.05798
-
MedREQAL: Examining Medical Knowledge Recall of Large Language Models via Question Answering 9 Jun 2024 · 0 repositories · arXiv:2406.05845
-
Text2VP: Generative AI for Visual Programming and Parametric Modeling 9 Jun 2024 · 0 repositories · arXiv:2407.07732
-
Critical Phase Transition in Large Language Models 8 Jun 2024 · 0 repositories · arXiv:2406.05335
-
Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts 8 Jun 2024 · 0 repositories · arXiv:2406.05569
-
MaTableGPT: GPT-based Table Data Extractor from Materials Science Literature 8 Jun 2024 · 0 repositories · arXiv:2406.05431
-
SelfDefend: LLMs Can Defend Themselves against Jailbreaking in a Practical Manner 8 Jun 2024 · 0 repositories · arXiv:2406.05498
-
BAMO at SemEval-2024 Task 9: BRAINTEASER: A Novel Task Defying Common Sense 7 Jun 2024 · 1 repository · arXiv:2406.04947
-
BERTs are Generative In-Context Learners 7 Jun 2024 · 1 repository · arXiv:2406.04823Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 1 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 13 unverified (of 26 harvested samples)
-
GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents 7 Jun 2024 · 2 repositories · arXiv:2406.06613Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Large Generative Graph Models 7 Jun 2024 · 0 repositories · arXiv:2406.05109
-
Low-Resource Cross-Lingual Summarization through Few-Shot Learning with Large Language Models 7 Jun 2024 · 0 repositories · arXiv:2406.04630
-
On Subjective Uncertainty Quantification and Calibration in Natural Language Generation 7 Jun 2024 · 1 repository · arXiv:2406.05213Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 14 harvested samples)
-
VTrans: Accelerating Transformer Compression with Variational Information Bottleneck based Pruning 7 Jun 2024 · 0 repositories · arXiv:2406.05276
-
A Survey on Medical Large Language Models: Technology, Application, Trustworthiness, and Future Directions 6 Jun 2024 · 0 repositories · arXiv:2406.03712
-
Do Language Models Understand Morality? Towards a Robust Detection of Moral Content 6 Jun 2024 · 1 repository · arXiv:2406.04143
-
HORAE: A Domain-Agnostic Language for Automated Service Regulation 6 Jun 2024 · 1 repository · arXiv:2406.06600
-
LLMEmbed: Rethinking Lightweight LLM's Genuine Function in Text Classification 6 Jun 2024 · 1 repository · arXiv:2406.03725
-
Simplified and Generalized Masked Diffusion for Discrete Data 6 Jun 2024 · 1 repository · arXiv:2406.04329Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 4 where Syntology's instrument failed) · 15 unverified (of 21 harvested samples)
-
Tox-BART: Leveraging Toxicity Attributes for Explanation Generation of Implicit Hate Speech 6 Jun 2024 · 1 repository · arXiv:2406.03953
-
Your Absorbing Discrete Diffusion Secretly Models the Conditional Distributions of Clean Data 6 Jun 2024 · 2 repositories · arXiv:2406.03736Syntology official (archive's flag): 13 ran · 13 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 1 honoured, 0 violated, 11 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Automating Turkish Educational Quiz Generation Using Large Language Models 5 Jun 2024 · 4 repositories · arXiv:2406.03397
-
Exact Conversion of In-Context Learning to Model Weights in Linearized-Attention Transformers 5 Jun 2024 · 0 repositories · arXiv:2406.02847
-
Exploring Multilingual Large Language Models for Enhanced TNM classification of Radiology Report in lung cancer staging 5 Jun 2024 · 0 repositories · arXiv:2406.06591
-
Missci: Reconstructing Fallacies in Misrepresented Science 5 Jun 2024 · 2 repositories · arXiv:2406.03181
-
RICo: Reddit ideological communities 5 Jun 2024 · 1 repository
-
StatBot.Swiss: Bilingual Open Data Exploration in Natural Language 5 Jun 2024 · 0 repositories · arXiv:2406.03170
-
The Good, the Bad, and the Hulk-like GPT: Analyzing Emotional Decisions of Large Language Models in Cooperation and Bargaining Games 5 Jun 2024 · 0 repositories · arXiv:2406.03299
-
Too Big to Fail: Larger Language Models are Disproportionately Resilient to Induction of Dementia-Related Linguistic Anomalies 5 Jun 2024 · 1 repository · arXiv:2406.02830
-
CheckEmbed: Effective Verification of LLM Solutions to Open-Ended Tasks 4 Jun 2024 · 1 repository · arXiv:2406.02524
-
Large Language Model-Enabled Multi-Agent Manufacturing Systems 4 Jun 2024 · 0 repositories · arXiv:2406.01893
-
OccamLLM: Fast and Exact Language Model Arithmetic in a Single Step 4 Jun 2024 · 0 repositories · arXiv:2406.06576
-
Randomized Geometric Algebra Methods for Convex Neural Networks 4 Jun 2024 · 1 repository · arXiv:2406.02806Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Annotation Guidelines-Based Knowledge Augmentation: Towards Enhancing Large Language Models for Educational Text Classification 3 Jun 2024 · 0 repositories · arXiv:2406.00954
-
In-Context Learning of Physical Properties: Few-Shot Adaptation to Out-of-Distribution Molecular Graphs 3 Jun 2024 · 0 repositories · arXiv:2406.01808
-
Luna: An Evaluation Foundation Model to Catch Language Model Hallucinations with High Accuracy and Low Cost 3 Jun 2024 · 0 repositories · arXiv:2406.00975
-
SemCoder: Training Code Language Models with Comprehensive Semantics Reasoning 3 Jun 2024 · 1 repository · arXiv:2406.01006Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples)
-
SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models 3 Jun 2024 · 1 repository · arXiv:2406.01584
-
Superhuman performance in urology board questions by an explainable large language model enabled for context integration of the European Association of Urology guidelines: the UroBot study 3 Jun 2024 · 0 repositories · arXiv:2406.01428
-
Unsupervised Distractor Generation via Large Language Model Distilling and Counterfactual Contrastive Decoding 3 Jun 2024 · 0 repositories · arXiv:2406.01306
-
Applying Fine-Tuned LLMs for Reducing Data Needs in Load Profile Analysis 2 Jun 2024 · 0 repositories · arXiv:2406.02479