Methods › General › Learning Rate Schedules › Linear Warmup With Cosine Annealing › Papers, page 9
Linear Warmup With Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,797 · with a code link: 1,655 · where Syntology ran a sample: 602 (490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (602 of 3,797 tagged: 490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument)
Page 9 of 38: papers 801 to 900 of 3,797, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Transformers and Large Language Models for Efficient Intrusion Detection Systems: A Comprehensive Survey 14 Aug 2024 · 0 repositories · arXiv:2408.07583
-
Evaluating Cultural Adaptability of a Large Language Model via Simulation of Synthetic Personas 13 Aug 2024 · 1 repository · arXiv:2408.06929
-
Generative AI for automatic topic labelling 13 Aug 2024 · 0 repositories · arXiv:2408.07003
-
Pragmatic inference of scalar implicature by LLMs 13 Aug 2024 · 0 repositories · arXiv:2408.06673
-
Kov: Transferable and Naturalistic Black-Box LLM Attacks using Markov Decision Processes and Tree Search 11 Aug 2024 · 1 repository · arXiv:2408.08899
-
PhishLang: A Real-Time, Fully Client-Side Phishing Detection Framework Using MobileBERT 11 Aug 2024 · 2 repositories · arXiv:2408.05667
-
Chain of Condition: Construct, Verify and Solve Conditions for Conditional Question Answering 10 Aug 2024 · 0 repositories · arXiv:2408.05442
-
Improving Whisper's Recognition Performance for Under-Represented Language Kazakh Leveraging Unpaired Speech and Text 10 Aug 2024 · 0 repositories · arXiv:2408.05554
-
COAST: Enhancing the Code Debugging Ability of LLMs through Communicative Agent Based Data Synthesis 9 Aug 2024 · 1 repository · arXiv:2408.05006
-
Examining the Behavior of LLM Architectures Within the Framework of Standardized National Exams in Brazil 9 Aug 2024 · 0 repositories · arXiv:2408.05035
-
From Text to Insight: Leveraging Large Language Models for Performance Evaluation in Management 9 Aug 2024 · 0 repositories · arXiv:2408.05328
-
Retrieval-augmented code completion for local projects using large language models 9 Aug 2024 · 0 repositories · arXiv:2408.05026
-
Transformer Explainer: Interactive Learning of Text-Generative Models 8 Aug 2024 · 1 repository · arXiv:2408.04619
-
Could ChatGPT get an Engineering Degree? Evaluating Higher Education Vulnerability to AI Assistants 7 Aug 2024 · 0 repositories · arXiv:2408.11841
-
Image-to-LaTeX Converter for Mathematical Formulas and Text 7 Aug 2024 · 1 repository · arXiv:2408.04015
-
Is Child-Directed Speech Effective Training Data for Language Models? 7 Aug 2024 · 1 repository · arXiv:2408.03617Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 2 pointer-only (licence)
-
Query3D: LLM-Powered Open-Vocabulary Scene Segmentation with Language Embedded 3D Gaussian 7 Aug 2024 · 1 repository · arXiv:2408.03516
-
SocFedGPT: Federated GPT-based Adaptive Content Filtering System Leveraging User Interactions in Social Networks 7 Aug 2024 · 0 repositories · arXiv:2408.05243
-
Data Poisoning in LLMs: Jailbreak-Tuning and Scaling Laws 6 Aug 2024 · 2 repositories · arXiv:2408.02946Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Evaluating the Translation Performance of Large Language Models Based on Euas-20 6 Aug 2024 · 0 repositories · arXiv:2408.03119
-
FLASH: Federated Learning-Based LLMs for Advanced Query Processing in Social Networks through RAG 6 Aug 2024 · 0 repositories · arXiv:2408.05242
-
The Use of Large Language Models (LLM) for Cyber Threat Intelligence (CTI) in Cybercrime Forums 6 Aug 2024 · 0 repositories · arXiv:2408.03354
-
TrafficGPT: An LLM Approach for Open-Set Encrypted Traffic Classification 6 Aug 2024 · 1 repository
-
Why Are My Prompts Leaked? Unraveling Prompt Extraction Threats in Customized Large Language Models 5 Aug 2024 · 1 repository · arXiv:2408.02416
-
XMainframe: A Large Language Model for Mainframe Modernization 5 Aug 2024 · 1 repository · arXiv:2408.04660
-
ML-EAT: A Multilevel Embedding Association Test for Interpretable and Transparent Social Science 4 Aug 2024 · 1 repository · arXiv:2408.01966
-
AdaCBM: An Adaptive Concept Bottleneck Model for Explainable and Accurate Diagnosis 4 Aug 2024 · 1 repository · arXiv:2408.02001
-
Effective Demonstration Annotation for In-Context Learning via Language Model-Based Determinantal Point Process 4 Aug 2024 · 0 repositories · arXiv:2408.02103
-
Leveraging Large Language Models with Chain-of-Thought and Prompt Engineering for Traffic Crash Severity Analysis and Inference 4 Aug 2024 · 0 repositories · arXiv:2408.04652
-
Advancing Mental Health Pre-Screening: A New Custom GPT for Psychological Distress Assessment 3 Aug 2024 · 0 repositories · arXiv:2408.01614
-
Efficient Solutions For An Intriguing Failure of LLMs: Long Context Window Does Not Mean LLMs Can Analyze Long Sequences Flawlessly 3 Aug 2024 · 0 repositories · arXiv:2408.01866
-
Building Trust in Mental Health Chatbots: Safety Metrics and LLM-Based Evaluation Tools 3 Aug 2024 · 0 repositories · arXiv:2408.04650
-
LLM as Runtime Error Handler: A Promising Pathway to Adaptive Self-Healing of Software Systems 2 Aug 2024 · 0 repositories · arXiv:2408.01055
-
High-Throughput Phenotyping of Clinical Text Using Large Language Models 2 Aug 2024 · 0 repositories · arXiv:2408.01214
-
Improving Retrieval-Augmented Generation in Medicine with Iterative Follow-up Questions 1 Aug 2024 · 1 repository · arXiv:2408.00727
-
AgentGen: Enhancing Planning Abilities for Large Language Model based Agent via Environment and Task Generation 1 Aug 2024 · 1 repository · arXiv:2408.00764Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples)
-
What comes after transformers? -- A selective survey connecting ideas in deep learning 1 Aug 2024 · 0 repositories · arXiv:2408.00386
-
Performance of Recent Large Language Models for a Low-Resourced Language 31 Jul 2024 · 0 repositories · arXiv:2407.21330
-
Improving Faithfulness of Large Language Models in Summarization via Sliding Generation and Self-Consistency 31 Jul 2024 · 0 repositories · arXiv:2407.21443
-
Generative Expressive Conversational Speech Synthesis 31 Jul 2024 · 1 repository · arXiv:2407.21491
-
Automated Software Vulnerability Static Code Analysis Using Generative Pre-Trained Transformer Models 31 Jul 2024 · 0 repositories · arXiv:2408.00197
-
Decomposed Prompting to Answer Questions on a Course Discussion Board 30 Jul 2024 · 1 repository · arXiv:2407.21170
-
BERT and LLMs-Based avGFP Brightness Prediction and Mutation Design 30 Jul 2024 · 0 repositories · arXiv:2407.20534
-
Breaking Agents: Compromising Autonomous LLM Agents Through Malfunction Amplification 30 Jul 2024 · 0 repositories · arXiv:2407.20859
-
Comparison of Large Language Models for Generating Contextually Relevant Questions 30 Jul 2024 · 1 repository · arXiv:2407.20578
-
AgEval: A Benchmark for Zero-Shot and Few-Shot Plant Stress Phenotyping with Multimodal LLMs 29 Jul 2024 · 0 repositories · arXiv:2407.19617
-
AutoScale: Scale-Aware Data Mixing for Pre-Training LLMs 29 Jul 2024 · 1 repository · arXiv:2407.20177Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Detecting and Understanding Vulnerabilities in Language Models via Mechanistic Interpretability 29 Jul 2024 · 1 repository · arXiv:2407.19842
-
Enhancing Code Translation in Language Models with Few-Shot Learning via Retrieval-Augmented Generation 29 Jul 2024 · 0 repositories · arXiv:2407.19619
-
AdaCoder: Adaptive Prompt Compression for Programmatic Visual Question Answering 28 Jul 2024 · 0 repositories · arXiv:2407.19410
-
Are LLMs Good Annotators for Discourse-level Event Relation Extraction? 28 Jul 2024 · 1 repository · arXiv:2407.19568
-
Is Generative AI an Existential Threat to Human Creatives? Insights from Financial Economics 28 Jul 2024 · 0 repositories · arXiv:2407.19586
-
Motamot: A Dataset for Revealing the Supremacy of Large Language Models over Transformer Models in Bengali Political Sentiment Analysis 28 Jul 2024 · 1 repository · arXiv:2407.19528
-
A Reliable Common-Sense Reasoning Socialbot Built Using LLMs and Goal-Directed ASP 26 Jul 2024 · 0 repositories · arXiv:2407.18498
-
Human-artificial intelligence teaming for scientific information extraction from data-driven additive manufacturing research using large language models 26 Jul 2024 · 0 repositories · arXiv:2407.18827
-
ClinicRealm: Re-evaluating Large Language Models with Conventional Machine Learning for Non-Generative Clinical Prediction Tasks 26 Jul 2024 · 1 repository · arXiv:2407.18525Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
TAGIFY: LLM-powered Tagging Interface for Improved Data Findability on OGD portals 26 Jul 2024 · 0 repositories · arXiv:2407.18764
-
Using Large Language Models for the Interpretation of Building Regulations 26 Jul 2024 · 0 repositories · arXiv:2407.21060
-
Closing the gap between open-source and commercial large language models for medical evidence summarization 25 Jul 2024 · 0 repositories · arXiv:2408.00588
-
Cost-effective Instruction Learning for Pathology Vision and Language Analysis 25 Jul 2024 · 1 repository · arXiv:2407.17734
-
PEFT-U: Parameter-Efficient Fine-Tuning for User Personalization 25 Jul 2024 · 1 repository · arXiv:2407.18078Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
PersonaGym: Evaluating Persona Agents and LLMs 25 Jul 2024 · 1 repository · arXiv:2407.18416
-
Bailicai: A Domain-Optimized Retrieval-Augmented Generation Framework for Medical Applications 24 Jul 2024 · 0 repositories · arXiv:2407.21055
-
Testing Large Language Models on Driving Theory Knowledge and Skills for Connected Autonomous Vehicles 24 Jul 2024 · 0 repositories · arXiv:2407.17211
-
Analyzing Polysemy Evolution Using Semantic Cells 23 Jul 2024 · 0 repositories · arXiv:2407.16110
-
Data Mixture Inference: What do BPE Tokenizers Reveal about their Training Data? 23 Jul 2024 · 1 repository · arXiv:2407.16607Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Enhancing LLM's Cognition via Structurization 23 Jul 2024 · 1 repository · arXiv:2407.16434Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Patched RTC: evaluating LLMs for diverse software development tasks 23 Jul 2024 · 1 repository · arXiv:2407.16557
-
Robust Privacy Amidst Innovation with Large Language Models Through a Critical Assessment of the Risks 23 Jul 2024 · 1 repository · arXiv:2407.16166
-
Impacts of Anthropomorphizing Large Language Models in Learning Environments 22 Jul 2024 · 0 repositories · arXiv:2408.03945
-
Imposter.AI: Adversarial Attacks with Hidden Intentions towards Aligned Large Language Models 22 Jul 2024 · 0 repositories · arXiv:2407.15399
-
Inverted Activations: Reducing Memory Footprint in Neural Network Training 22 Jul 2024 · 1 repository · arXiv:2407.15545
-
MMInstruct: A High-Quality Multi-Modal Instruction Tuning Dataset with Extensive Diversity 22 Jul 2024 · 1 repository · arXiv:2407.15838Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
RadioRAG: Factual large language models for enhanced diagnostics in radiology using online retrieval augmented generation 22 Jul 2024 · 1 repository · arXiv:2407.15621
-
Stretching Each Dollar: Diffusion Training from Scratch on a Micro-Budget 22 Jul 2024 · 1 repository · arXiv:2407.15811Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Unlocking the Potential: Benchmarking Large Language Models in Water Engineering and Research 22 Jul 2024 · 0 repositories · arXiv:2407.21045
-
Decoding Multilingual Moral Preferences: Unveiling LLM's Biases Through the Moral Machine Experiment 21 Jul 2024 · 1 repository · arXiv:2407.15184
-
SQLfuse: Enhancing Text-to-SQL Performance through Comprehensive LLM Synergy 19 Jul 2024 · 0 repositories · arXiv:2407.14568
-
Scalable Exploration via Ensemble++ 18 Jul 2024 · 2 repositories · arXiv:2407.13195Syntology official (archive's flag): 5 ran · 5 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Can Open-Source LLMs Compete with Commercial Models? Exploring the Few-Shot Performance of Current GPT Models in Biomedical Tasks 18 Jul 2024 · 1 repository · arXiv:2407.13511
-
Evaluating Large Language Models for Anxiety and Depression Classification using Counseling and Psychotherapy Transcripts 18 Jul 2024 · 1 repository · arXiv:2407.13228
-
How Reliable are LLMs as Knowledge Bases? Re-thinking Facutality and Consistency 18 Jul 2024 · 0 repositories · arXiv:2407.13578
-
Learning-From-Mistakes Prompting for Indigenous Language Translation 18 Jul 2024 · 0 repositories · arXiv:2407.13343
-
PRAGyan -- Connecting the Dots in Tweets 18 Jul 2024 · 0 repositories · arXiv:2407.13909
-
Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction 18 Jul 2024 · 1 repository · arXiv:2407.13943Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Beyond Binary: Multiclass Paraphasia Detection with Generative Pretrained Transformers and End-to-End Models 16 Jul 2024 · 0 repositories · arXiv:2407.11345
-
ChatBCG: Can AI Read Your Slide Deck? 16 Jul 2024 · 0 repositories · arXiv:2407.12875
-
Does Refusal Training in LLMs Generalize to the Past Tense? 16 Jul 2024 · 1 repository · arXiv:2407.11969Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
GPT Assisted Annotation of Rhetorical and Linguistic Features for Interpretable Propaganda Technique Detection in News Text 16 Jul 2024 · 0 repositories · arXiv:2407.11827
-
LaMI-DETR: Open-Vocabulary Detection with Language Model Instruction 16 Jul 2024 · 1 repository · arXiv:2407.11335Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Large Language Models as Misleading Assistants in Conversation 16 Jul 2024 · 0 repositories · arXiv:2407.11789
-
Large Visual-Language Models Are Also Good Classifiers: A Study of In-Context Multimodal Fake News Detection 16 Jul 2024 · 0 repositories · arXiv:2407.12879
-
Representation Bias in Political Sample Simulations with Large Language Models 16 Jul 2024 · 0 repositories · arXiv:2407.11409
-
ReFeR: Improving Evaluation and Reasoning through Hierarchy of Models 16 Jul 2024 · 0 repositories · arXiv:2407.12877
-
Trust No Bot: Discovering Personal Disclosures in Human-LLM Conversations in the Wild 16 Jul 2024 · 1 repository · arXiv:2407.11438
-
CodeV: Empowering LLMs with HDL Generation through Multi-Level Summarization 15 Jul 2024 · 0 repositories · arXiv:2407.10424
-
Leveraging LLM-Respondents for Item Evaluation: a Psychometric Analysis 15 Jul 2024 · 0 repositories · arXiv:2407.10899
-
Making New Connections: LLMs as Puzzle Generators for The New York Times' Connections Word Game 15 Jul 2024 · 0 repositories · arXiv:2407.11240
-
Mechanistic interpretability of large language models with applications to the financial services industry 15 Jul 2024 · 0 repositories · arXiv:2407.11215
-
MetaLLM: A High-performant and Cost-efficient Dynamic Framework for Wrapping LLMs 15 Jul 2024 · 1 repository · arXiv:2407.10834Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)