Methods › General › Learning Rate Schedules › Linear Warmup With Cosine Annealing › Papers, page 5
Linear Warmup With Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,797 · with a code link: 1,655 · where Syntology ran a sample: 602 (490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (602 of 3,797 tagged: 490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument)
Page 5 of 38: papers 401 to 500 of 3,797, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Auto-Generating Earnings Report Analysis via a Financial-Augmented LLM 11 Dec 2024 · 0 repositories · arXiv:2412.08179
-
Exploiting the Index Gradients for Optimization-Based Jailbreaking on Large Language Models 11 Dec 2024 · 1 repository · arXiv:2412.08615
-
GraphTool-Instruction: Revolutionizing Graph Reasoning in LLMs through Decomposed Subtask Instruction 11 Dec 2024 · 0 repositories · arXiv:2412.12152
-
Imitate Before Detect: Aligning Machine Stylistic Preference for Machine-Revised Text Detection 11 Dec 2024 · 0 repositories · arXiv:2412.10432
-
Large Language Models Still Face Challenges in Multi-Hop Reasoning with External Knowledge 11 Dec 2024 · 0 repositories · arXiv:2412.08317
-
A Causal World Model Underlying Next Token Prediction: Exploring GPT in a Controlled Environment 10 Dec 2024 · 1 repository · arXiv:2412.07446
-
GPT-2 Through the Lens of Vector Symbolic Architectures 10 Dec 2024 · 0 repositories · arXiv:2412.07947
-
IntellectSeeker: A Personalized Literature Management System with the Probabilistic Model and Large Language Model 10 Dec 2024 · 1 repository · arXiv:2412.07213
-
RAG-based Question Answering over Heterogeneous Data and Text 10 Dec 2024 · 0 repositories · arXiv:2412.07420
-
Superficial Consciousness Hypothesis for Autoregressive Transformers 10 Dec 2024 · 1 repository · arXiv:2412.07278
-
Towards Predictive Communication with Brain-Computer Interfaces integrating Large Language Models 10 Dec 2024 · 0 repositories · arXiv:2412.07355
-
BatchTopK Sparse Autoencoders 9 Dec 2024 · 2 repositories · arXiv:2412.06410Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Exploring Memorization and Copyright Violation in Frontier LLMs: A Study of the New York Times v. OpenAI 2023 Lawsuit 9 Dec 2024 · 0 repositories · arXiv:2412.06370
-
Optimizing Multi-Task Learning for Enhanced Performance in Large Language Models 9 Dec 2024 · 0 repositories · arXiv:2412.06249
-
The Rosetta Paradox: Domain-Specific Performance Inversions in Large Language Models 9 Dec 2024 · 0 repositories · arXiv:2412.17821
-
M³-20M: A Large-Scale Multi-Modal Molecule Dataset for AI-driven Drug Design and Discovery 8 Dec 2024 · 1 repository · arXiv:2412.06847
-
CharacterBox: Evaluating the Role-Playing Capabilities of LLMs in Text-Based Virtual Worlds 7 Dec 2024 · 1 repository · arXiv:2412.05631Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
Can the Rookies Cut the Tough Cookie? Exploring the Use of LLMs for SQL Equivalence Checking 7 Dec 2024 · 0 repositories · arXiv:2412.05561
-
PrivAgent: Agentic-based Red-teaming for LLM Privacy Leakage 7 Dec 2024 · 1 repository · arXiv:2412.05734
-
100% Elimination of Hallucinations on RAGTruth for GPT-4 and GPT-3.5 Turbo 6 Dec 2024 · 0 repositories · arXiv:2412.05223
-
Are Frontier Large Language Models Suitable for Q&A in Science Centres? 6 Dec 2024 · 0 repositories · arXiv:2412.05200
-
QueEn: A Large Language Model for Quechua-English Translation 6 Dec 2024 · 0 repositories · arXiv:2412.05184
-
How Good is ChatGPT in Giving Adaptive Guidance Using Knowledge Graphs in E-Learning Environments? 5 Dec 2024 · 0 repositories · arXiv:2412.03856
-
Controlling the Mutation in Large Language Models for the Efficient Evolution of Algorithms 4 Dec 2024 · 0 repositories · arXiv:2412.03250
-
Compressing KV Cache for Long-Context LLM Inference with Inter-Layer Attention Similarity 3 Dec 2024 · 0 repositories · arXiv:2412.02252
-
DP-2Stage: Adapting Language Models as Differentially Private Tabular Data Generators 3 Dec 2024 · 1 repository · arXiv:2412.02467
-
Flattering to Deceive: The Impact of Sycophantic Behavior on User Trust in Large Language Model 3 Dec 2024 · 0 repositories · arXiv:2412.02802
-
Impact of Data Snooping on Deep Learning Models for Locating Vulnerabilities in Lifted Code 3 Dec 2024 · 0 repositories · arXiv:2412.02048
-
The Asymptotic Behavior of Attention in Transformers 3 Dec 2024 · 0 repositories · arXiv:2412.02682
-
Su-RoBERTa: A Semi-supervised Approach to Predicting Suicide Risk through Social Media using Base Language Models 2 Dec 2024 · 0 repositories · arXiv:2412.01353
-
The Promise and Peril of Generative AI: Evidence from GPT-4 as Sell-Side Analysts 2 Dec 2024 · 0 repositories · arXiv:2412.01069
-
Tokenizing 3D Molecule Structure with Quantized Spherical Coordinates 2 Dec 2024 · 0 repositories · arXiv:2412.01564
-
A Comprehensive Guide to Explainable AI: From Classical Models to LLMs 1 Dec 2024 · 1 repository · arXiv:2412.00800
-
EventGPT: Event Stream Understanding with Multimodal Large Language Models 1 Dec 2024 · 0 repositories · arXiv:2412.00832
-
CDEMapper: Enhancing NIH Common Data Element Normalization using Large Language Models 30 Nov 2024 · 0 repositories · arXiv:2412.00491
-
Cognitive Biases in Large Language Models: A Survey and Mitigation Experiments 30 Nov 2024 · 0 repositories · arXiv:2412.00323
-
Empowering the Deaf and Hard of Hearing Community: Enhancing Video Captions Using Large Language Models 30 Nov 2024 · 0 repositories · arXiv:2412.00342
-
Forma mentis networks predict creativity ratings of short texts via interpretable artificial intelligence in human and GPT-simulated raters 30 Nov 2024 · 0 repositories · arXiv:2412.00530
-
Automatic Prompt Generation and Grounding Object Detection for Zero-Shot Image Anomaly Detection 28 Nov 2024 · 0 repositories · arXiv:2411.19220
-
Beautimeter: Harnessing GPT for Assessing Architectural and Urban Beauty based on the 15 Properties of Living Structure 28 Nov 2024 · 0 repositories · arXiv:2411.19094
-
DENIAHL: In-Context Features Influence LLM Needle-In-A-Haystack Abilities 28 Nov 2024 · 1 repository · arXiv:2411.19360
-
Habit Coach: Customising RAG-based chatbots to support behavior change 28 Nov 2024 · 0 repositories · arXiv:2411.19229
-
SmartLLMSentry: A Comprehensive LLM Based Smart Contract Vulnerability Detection Framework 28 Nov 2024 · 0 repositories · arXiv:2411.19234
-
The Impact of Example Selection in Few-Shot Prompting on Automated Essay Scoring Using GPT Models 28 Nov 2024 · 0 repositories · arXiv:2411.18924
-
Automated Literature Review Using NLP Techniques and LLM-Based Retrieval-Augmented Generation 27 Nov 2024 · 0 repositories · arXiv:2411.18583
-
Can bidirectional encoder become the ultimate winner for downstream applications of foundation models? 27 Nov 2024 · 0 repositories · arXiv:2411.18021
-
ChatGPT as speechwriter for the French presidents 27 Nov 2024 · 0 repositories · arXiv:2411.18382
-
DRS: Deep Question Reformulation With Structured Output 27 Nov 2024 · 1 repository · arXiv:2411.17993
-
Streamlining Prediction in Bayesian Deep Learning 27 Nov 2024 · 1 repository · arXiv:2411.18425Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples)
-
Training and Evaluating Language Models with Template-based Data Generation 27 Nov 2024 · 1 repository · arXiv:2411.18104Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Advancing Content Moderation: Evaluating Large Language Models for Detecting Sensitive Content Across Text, Images, and Videos 26 Nov 2024 · 0 repositories · arXiv:2411.17123
-
Can artificial intelligence predict clinical trial outcomes? 26 Nov 2024 · 0 repositories · arXiv:2411.17595
-
CLOVER: Cross-Layer Orthogonal Vectors Pruning and Fine-Tuning 26 Nov 2024 · 1 repository · arXiv:2411.17426
-
Distributed Sign Momentum with Local Steps for Training Transformers 26 Nov 2024 · 1 repository · arXiv:2411.17866
-
"Give me the code" -- Log Analysis of First-Year CS Students' Interactions With GPT 26 Nov 2024 · 0 repositories · arXiv:2411.17855
-
On Limitations of LLM as Annotator for Low Resource Languages 26 Nov 2024 · 0 repositories · arXiv:2411.17637
-
Pretrained LLM Adapted with LoRA as a Decision Transformer for Offline RL in Quantitative Trading 26 Nov 2024 · 1 repository · arXiv:2411.17900
-
Adaptive Circuit Behavior and Generalization in Mechanistic Interpretability 25 Nov 2024 · 0 repositories · arXiv:2411.16105
-
Are Transformers Truly Foundational for Robotics? 25 Nov 2024 · 0 repositories · arXiv:2411.16917
-
Can AI grade your essays? A comparative analysis of large language models and teacher ratings in multidimensional essay scoring 25 Nov 2024 · 0 repositories · arXiv:2411.16337
-
Fine-Tuning LLMs with Noisy Data for Political Argument Generation and Post Guidance 25 Nov 2024 · 0 repositories · arXiv:2411.16813
-
MarketGPT: Developing a Pre-trained transformer (GPT) for Modeling Financial Time Series 25 Nov 2024 · 1 repository · arXiv:2411.16585Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Predictive Power of LLMs in Financial Markets 25 Nov 2024 · 0 repositories · arXiv:2411.16569
-
Development of Pre-Trained Transformer-based Models for the Nepali Language 24 Nov 2024 · 0 repositories · arXiv:2411.15734
-
"All that Glitters": Approaches to Evaluations with Unreliable Model and Human Annotations 23 Nov 2024 · 1 repository · arXiv:2411.15634
-
ChatBCI: A P300 Speller BCI Leveraging Large Language Models for Improved Sentence Composition in Realistic Scenarios 23 Nov 2024 · 0 repositories · arXiv:2411.15395
-
Improving Next Tokens via Second-Last Predictions with Generate and Refine 23 Nov 2024 · 0 repositories · arXiv:2411.15661
-
Comparative Analysis of Pooling Mechanisms in LLMs: A Sentiment Analysis Perspective 22 Nov 2024 · 0 repositories · arXiv:2411.14654
-
Assessment of LLM Responses to End-user Security Questions 21 Nov 2024 · 0 repositories · arXiv:2411.14571
-
Evaluating the Robustness of Analogical Reasoning in Large Language Models 21 Nov 2024 · 1 repository · arXiv:2411.14215Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
AI-Driven Agents with Prompts Designed for High Agreeableness Increase the Likelihood of Being Mistaken for a Human in the Turing Test 20 Nov 2024 · 0 repositories · arXiv:2411.13749
-
Combining Autoregressive and Autoencoder Language Models for Text Classification 20 Nov 2024 · 1 repository · arXiv:2411.13282
-
Exploring Large Language Models for Climate Forecasting 20 Nov 2024 · 0 repositories · arXiv:2411.13724
-
A Combined Encoder and Transformer Approach for Coherent and High-Quality Text Generation 19 Nov 2024 · 0 repositories · arXiv:2411.12157
-
Leveraging Virtual Reality and AI Tutoring for Language Learning: A Case Study of a Virtual Campus Environment with OpenAI GPT Integration with Unity 3D 19 Nov 2024 · 0 repositories · arXiv:2411.12619
-
Can Open-source LLMs Enhance Data Synthesis for Toxic Detection?: An Experimental Study 18 Nov 2024 · 0 repositories · arXiv:2411.15175
-
Chapter 7 Review of Data-Driven Generative AI Models for Knowledge Extraction from Scientific Literature in Healthcare 18 Nov 2024 · 0 repositories · arXiv:2411.11635
-
CNMBERT: A Model for Converting Hanyu Pinyin Abbreviations to Chinese Characters 18 Nov 2024 · 1 repository · arXiv:2411.11770
-
PerfCodeGen: Improving Performance of LLM Generated Code with Execution Feedback 18 Nov 2024 · 1 repository · arXiv:2412.03578Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Understanding Student Sentiment on Mental Health Support in Colleges Using Large Language Models 18 Nov 2024 · 0 repositories · arXiv:2412.04326
-
VersaTune: An Efficient Data Composition Framework for Training Multi-Capability LLMs 18 Nov 2024 · 1 repository · arXiv:2411.11266
-
Does Prompt Formatting Have Any Impact on LLM Performance? 15 Nov 2024 · 0 repositories · arXiv:2411.10541
-
MARS: Unleashing the Power of Variance Reduction for Training Large Models 15 Nov 2024 · 2 repositories · arXiv:2411.10438Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Prompting and Fine-tuning Large Language Models for Automated Code Review Comment Generation 15 Nov 2024 · 0 repositories · arXiv:2411.10129
-
Take Package as Language: Anomaly Detection Using Transformer 15 Nov 2024 · 0 repositories · arXiv:2412.04473
-
BabyLM Challenge: Exploring the Effect of Variation Sets on Language Model Training Efficiency 14 Nov 2024 · 0 repositories · arXiv:2411.09587
-
Beyond Static Tools: Evaluating Large Language Models for Cryptographic Misuse Detection 14 Nov 2024 · 0 repositories · arXiv:2411.09772
-
HateGPT: Unleashing GPT-3.5 Turbo to Combat Hate Speech on X 14 Nov 2024 · 0 repositories · arXiv:2411.09214
-
LLMStinger: Jailbreaking LLMs using RL fine-tuned LLMs 13 Nov 2024 · 0 repositories · arXiv:2411.08862
-
Responsible AI in Construction Safety: Systematic Evaluation of Large Language Models and Prompt Engineering 13 Nov 2024 · 0 repositories · arXiv:2411.08320
-
VALTEST: Automated Validation of Language Model Generated Test Cases 13 Nov 2024 · 0 repositories · arXiv:2411.08254
-
Evaluating ChatGPT-3.5 Efficiency in Solving Coding Problems of Different Complexity Levels: An Empirical Analysis 12 Nov 2024 · 1 repository · arXiv:2411.07529
-
Fair Summarization: Bridging Quality and Diversity in Extractive Summaries 12 Nov 2024 · 1 repository · arXiv:2411.07521Syntology official (archive's flag): 3 ran · 4 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
LLM App Squatting and Cloning 12 Nov 2024 · 0 repositories · arXiv:2411.07518
-
Ambient AI Scribing Support: Comparing the Performance of Specialized AI Agentic Architecture to Leading Foundational Models 11 Nov 2024 · 0 repositories · arXiv:2411.06713
-
Autonomous Droplet Microfluidic Design Framework with Large Language Models 11 Nov 2024 · 1 repository · arXiv:2411.06691
-
Cancer-Answer: Empowering Cancer Care with Advanced Large Language Models 11 Nov 2024 · 0 repositories · arXiv:2411.06946
-
Explore the Reasoning Capability of LLMs in the Chess Testbed 11 Nov 2024 · 0 repositories · arXiv:2411.06655
-
On Active Privacy Auditing in Supervised Fine-tuning for White-Box Language Models 11 Nov 2024 · 0 repositories · arXiv:2411.07070
-
SPARTAN: A Sparse Transformer Learning Local Causation 11 Nov 2024 · 0 repositories · arXiv:2411.06890