Methods › General › Learning Rate Schedules › Linear Warmup With Cosine Annealing › Papers, page 22
Linear Warmup With Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,797 · with a code link: 1,655 · where Syntology ran a sample: 602 (490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (602 of 3,797 tagged: 490 with a run with no instrument failure, 112 where every run was a failure of Syntology's instrument)
Page 22 of 38: papers 2,101 to 2,200 of 3,797, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Memory Gym: Towards Endless Tasks to Benchmark Memory Capabilities of Agents 29 Sep 2023 · 1 repository · arXiv:2309.17207Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Revolutionizing Mobile Interaction: Enabling a 3 Billion Parameter GPT LLM on Mobile 29 Sep 2023 · 0 repositories · arXiv:2310.01434
-
Split and Merge: Aligning Position Biases in LLM-based Evaluators 29 Sep 2023 · 0 repositories · arXiv:2310.01432
-
Training and inference of large language models using 8-bit floating point 29 Sep 2023 · 0 repositories · arXiv:2309.17224
-
AE-GPT: Using Large Language Models to Extract Adverse Events from Surveillance Reports-A Use Case with Influenza Vaccine Adverse Events 28 Sep 2023 · 0 repositories · arXiv:2309.16150
-
GPT-Fathom: Benchmarking Large Language Models to Decipher the Evolutionary Path towards GPT-4 and Beyond 28 Sep 2023 · 1 repository · arXiv:2309.16583Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples)
-
Large Language Model Soft Ideologization via AI-Self-Consciousness 28 Sep 2023 · 0 repositories · arXiv:2309.16167
-
Stress Testing Chain-of-Thought Prompting for Large Language Models 28 Sep 2023 · 0 repositories · arXiv:2309.16621
-
MindGPT: Interpreting What You See with Non-invasive Brain Recordings 27 Sep 2023 · 1 repository · arXiv:2309.15729Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
NLPBench: Evaluating Large Language Models on Solving NLP Problems 27 Sep 2023 · 1 repository · arXiv:2309.15630
-
Legal Question-Answering in the Indian Context: Efficacy, Challenges, and Potential of Modern AI Models 26 Sep 2023 · 0 repositories · arXiv:2309.14735
-
Exploring Small Language Models with Prompt-Learning Paradigm for Efficient Domain-Specific Text Classification 26 Sep 2023 · 0 repositories · arXiv:2309.14779
-
How to Catch an AI Liar: Lie Detection in Black-Box LLMs by Asking Unrelated Questions 26 Sep 2023 · 1 repository · arXiv:2309.15840Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
RankVicuna: Zero-Shot Listwise Document Reranking with Open-Source Large Language Models 26 Sep 2023 · 3 repositories · arXiv:2309.15088
-
Supersonic: Learning to Generate Source Code Optimizations in C/C++ 26 Sep 2023 · 1 repository · arXiv:2309.14846
-
Evaluating Cognitive Maps and Planning in Large Language Models with CogEval 25 Sep 2023 · 0 repositories · arXiv:2309.15129
-
LogGPT: Log Anomaly Detection via GPT 25 Sep 2023 · 1 repository · arXiv:2309.14482
-
Watch Your Language: Investigating Content Moderation with Large Language Models 25 Sep 2023 · 0 repositories · arXiv:2309.14517
-
Does the "most sinfully decadent cake ever" taste good? Answering Yes/No Questions from Figurative Contexts 24 Sep 2023 · 0 repositories · arXiv:2309.13748
-
Seeing Is Not Always Believing: Invisible Collision Attack and Defence on Pre-Trained Models 24 Sep 2023 · 1 repository · arXiv:2309.13579
-
A Chat About Boring Problems: Studying GPT-based text normalization 23 Sep 2023 · 0 repositories · arXiv:2309.13426
-
Probing the Moral Development of Large Language Models through Defining Issues Test 23 Sep 2023 · 0 repositories · arXiv:2309.13356
-
AMPLIFY:Attention-based Mixup for Performance Improvement and Label Smoothing in Transformer 22 Sep 2023 · 1 repository · arXiv:2309.12689
-
BenLLMEval: A Comprehensive Evaluation into the Potentials and Pitfalls of Large Language Models on Bengali NLP 22 Sep 2023 · 0 repositories · arXiv:2309.13173
-
Contextual Emotion Estimation from Image Captions 22 Sep 2023 · 0 repositories · arXiv:2309.13136
-
Investigating Large Language Models and Control Mechanisms to Improve Text Readability of Biomedical Abstracts 22 Sep 2023 · 1 repository · arXiv:2309.13202
-
Large Language Models Are Also Good Prototypical Commonsense Reasoners 22 Sep 2023 · 0 repositories · arXiv:2309.13165
-
SPION: Layer-Wise Sparse Training of Transformer via Convolutional Flood Filling 22 Sep 2023 · 0 repositories · arXiv:2309.12578
-
Goal-Oriented Prompt Attack and Safety Evaluation for LLMs 21 Sep 2023 · 2 repositories · arXiv:2309.11830
-
Bad Actor, Good Advisor: Exploring the Role of Large Language Models in Fake News Detection 21 Sep 2023 · 1 repository · arXiv:2309.12247
-
Constraints First: A New MDD-based Model to Generate Sentences Under Constraints 21 Sep 2023 · 0 repositories · arXiv:2309.12415
-
MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models 21 Sep 2023 · 1 repository · arXiv:2309.12284Syntology official (archive's flag): 15 ran · 15 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 14 where Syntology's instrument failed) · 7 unverified (of 22 harvested samples)
-
Random-Access Infinite Context Length for Transformers 21 Sep 2023 · 1 repository
-
TART: A plug-and-play Transformer module for task-agnostic reasoning 21 Sep 2023 · 1 repository
-
The Cambridge Law Corpus: A Dataset for Legal AI Research 21 Sep 2023 · 0 repositories · arXiv:2309.12269
-
The Reversal Curse: LLMs trained on "A is B" fail to learn "B is A" 21 Sep 2023 · 2 repositories · arXiv:2309.12288Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
TOA: Task-oriented Active VQA 21 Sep 2023 · 0 repositories
-
A Paradigm Shift in Machine Translation: Boosting Translation Performance of Large Language Models 20 Sep 2023 · 1 repository · arXiv:2309.11674Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Controlled Generation with Prompt Insertion for Natural Language Explanations in Grammatical Error Correction 20 Sep 2023 · 1 repository · arXiv:2309.11439
-
Design of Chain-of-Thought in Math Problem Solving 20 Sep 2023 · 1 repository · arXiv:2309.11054
-
Fictional Worlds, Real Connections: Developing Community Storytelling Social Chatbots through LLMs 20 Sep 2023 · 0 repositories · arXiv:2309.11478
-
Generative AI in Mafia-like Game Simulation 20 Sep 2023 · 0 repositories · arXiv:2309.11672
-
Safurai 001: New Qualitative Approach for Code LLM Evaluation 20 Sep 2023 · 1 repository · arXiv:2309.11385
-
Sequence-to-Sequence Spanish Pre-trained Language Models 20 Sep 2023 · 1 repository · arXiv:2309.11259
-
The Languini Kitchen: Enabling Language Modelling Research at Different Scales of Compute 20 Sep 2023 · 1 repository · arXiv:2309.11197
-
Language as the Medium: Multimodal Video Classification through text only 19 Sep 2023 · 0 repositories · arXiv:2309.10783
-
Rigorously Assessing Natural Language Explanations of Neurons 19 Sep 2023 · 0 repositories · arXiv:2309.10312
-
Writer-Defined AI Personas for On-Demand Feedback Generation 19 Sep 2023 · 0 repositories · arXiv:2309.10433
-
Evaluation of GPT-3 for Anti-Cancer Drug Sensitivity Prediction 18 Sep 2023 · 0 repositories · arXiv:2309.10016
-
RECAP: Retrieval-Augmented Audio Captioning 18 Sep 2023 · 1 repository · arXiv:2309.09836Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Towards Ontology Construction with Language Models 18 Sep 2023 · 0 repositories · arXiv:2309.09898
-
Contrastive Decoding Improves Reasoning in Large Language Models 17 Sep 2023 · 0 repositories · arXiv:2309.09117
-
Do Large GPT Models Discover Moral Dimensions in Language Representations? A Topological Study Of Sentence Embeddings 17 Sep 2023 · 0 repositories · arXiv:2309.09397
-
From Cooking Recipes to Robot Task Trees -- Improving Planning Correctness and Task Efficiency by Leveraging LLMs with a Knowledge Network 17 Sep 2023 · 0 repositories · arXiv:2309.09181
-
Decoder-only Architecture for Speech Recognition with CTC Prompts and Text Data Augmentation 16 Sep 2023 · 0 repositories · arXiv:2309.08876
-
Struc-Bench: Are Large Language Models Really Good at Generating Complex Structured Data? 16 Sep 2023 · 1 repository · arXiv:2309.08963
-
A Modern Turkish Poet: Fine-Tuned GPT-2 15 Sep 2023 · 1 repository
-
Advancing the Evaluation of Traditional Chinese Language Models: Towards a Comprehensive Benchmark Suite 15 Sep 2023 · 1 repository · arXiv:2309.08448
-
Indian-BhED: A Dataset for Measuring India-Centric Biases in Large Language Models 15 Sep 2023 · 1 repository · arXiv:2309.08573
-
EvoPrompt: Connecting LLMs with Evolutionary Algorithms Yields Powerful Prompt Optimizers 15 Sep 2023 · 2 repositories · arXiv:2309.08532Syntology 11 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 8 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples)
-
CoCA: Fusing Position Embedding with Collinear Constrained Attention in Transformers for Long Context Window Extending 15 Sep 2023 · 1 repository · arXiv:2309.08646
-
GPT-Lab: Next Generation Of Optimal Chemistry Discovery By GPT Driven Robotic Lab 15 Sep 2023 · 0 repositories · arXiv:2309.16721
-
ICLEF: In-Context Learning with Expert Feedback for Explainable Style Transfer 15 Sep 2023 · 1 repository · arXiv:2309.08583
-
Large Language Models for Failure Mode Classification: An Investigation 15 Sep 2023 · 1 repository · arXiv:2309.08181
-
An Empirical Evaluation of Prompting Strategies for Large Language Models in Zero-Shot Clinical Natural Language Processing 14 Sep 2023 · 0 repositories · arXiv:2309.08008
-
Assessing the nature of large language models: A caution against anthropocentrism 14 Sep 2023 · 0 repositories · arXiv:2309.07683
-
ChatGPT MT: Competitive for High- (but not Low-) Resource Languages 14 Sep 2023 · 2 repositories · arXiv:2309.07423Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Two Timin': Repairing Smart Contracts With A Two-Layered Approach 14 Sep 2023 · 0 repositories · arXiv:2309.07841
-
Large Language Models Can Infer Psychological Dispositions of Social Media Users 13 Sep 2023 · 0 repositories · arXiv:2309.08631
-
Traveling Words: A Geometric Interpretation of Transformers 13 Sep 2023 · 1 repository · arXiv:2309.07315
-
Circuit Breaking: Removing Model Behaviors with Targeted Ablation 12 Sep 2023 · 1 repository · arXiv:2309.05973Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Characterizing Latent Perspectives of Media Houses Towards Public Figures 12 Sep 2023 · 0 repositories · arXiv:2309.06112
-
Comparing Llama-2 and GPT-3 LLMs for HPC kernels generation 12 Sep 2023 · 0 repositories · arXiv:2309.07103
-
Exploring Large Language Models for Ontology Alignment 12 Sep 2023 · 1 repository · arXiv:2309.07172
-
Strategic Behavior of Large Language Models: Game Structure vs. Contextual Framing 12 Sep 2023 · 0 repositories · arXiv:2309.05898
-
The Moral Machine Experiment on Large Language Models 12 Sep 2023 · 1 repository · arXiv:2309.05958
-
Unveiling the potential of large language models in generating semantic and cross-language clones 12 Sep 2023 · 0 repositories · arXiv:2309.06424
-
Black-Box Analysis: GPTs Across Time in Legal Textual Entailment Task 11 Sep 2023 · 0 repositories · arXiv:2309.05501
-
Memory Injections: Correcting Multi-Hop Reasoning Failures during Inference in Transformer-Based Language Models 11 Sep 2023 · 1 repository · arXiv:2309.05605Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
SparseSwin: Swin Transformer with Sparse Transformer Block 11 Sep 2023 · 1 repository · arXiv:2309.05224
-
Zero-shot Learning with Minimum Instruction to Extract Social Determinants and Family History from Clinical Notes using GPT Model 11 Sep 2023 · 0 repositories · arXiv:2309.05475
-
Implementing Learning Principles with a Personal AI Tutor: A Case Study 10 Sep 2023 · 0 repositories · arXiv:2309.13060
-
Can NLP Models 'Identify', 'Distinguish', and 'Justify' Questions that Don't have a Definitive Answer? 8 Sep 2023 · 0 repositories · arXiv:2309.04635
-
Context-Aware Prompt Tuning for Vision-Language Model with Dual-Alignment 8 Sep 2023 · 0 repositories · arXiv:2309.04158
-
Evaluating ChatGPT as a Recommender System: A Rigorous Approach 7 Sep 2023 · 1 repository · arXiv:2309.03613
-
Supervised Learning and Large Language Model Benchmarks on Mental Health Datasets: Cognitive Distortions and Suicidal Risks in Chinese Social Media 7 Sep 2023 · 2 repositories · arXiv:2309.03564
-
FLM-101B: An Open LLM and How to Train It with $100K Budget 7 Sep 2023 · 0 repositories · arXiv:2309.03852
-
Zero-Shot Audio Captioning via Audibility Guidance 7 Sep 2023 · 0 repositories · arXiv:2309.03884
-
HAE-RAE Bench: Evaluation of Korean Knowledge in Language Models 6 Sep 2023 · 1 repository · arXiv:2309.02706
-
CodeApex: A Bilingual Programming Evaluation Benchmark for Large Language Models 5 Sep 2023 · 1 repository · arXiv:2309.01940
-
Do You Trust ChatGPT? -- Perceived Credibility of Human and AI-Generated Content 5 Sep 2023 · 0 repositories · arXiv:2309.02524
-
Do androids dream of fictional references? A bibliographic dialogue with ChatGPT3.5 4 Sep 2023 · 0 repositories · arXiv:2312.00789
-
Prompting or Fine-tuning? A Comparative Study of Large Language Models for Taxonomy Construction 4 Sep 2023 · 1 repository · arXiv:2309.01715
-
Saturn: An Optimized Data System for Large Model Deep Learning Workloads 3 Sep 2023 · 1 repository · arXiv:2309.01226
-
Large Language Models for Semantic Monitoring of Corporate Disclosures: A Case Study on Korea's Top 50 KOSPI Companies 1 Sep 2023 · 0 repositories · arXiv:2309.00208
-
Publicly Shareable Clinical Large Language Model Built on Synthetic Clinical Notes 1 Sep 2023 · 1 repository · arXiv:2309.00237Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Taken out of context: On measuring situational awareness in LLMs 1 Sep 2023 · 1 repository · arXiv:2309.00667Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Why do universal adversarial attacks work on large language models?: Geometry might be the answer 1 Sep 2023 · 0 repositories · arXiv:2309.00254
-
BioCoder: A Benchmark for Bioinformatics Code Generation with Large Language Models 31 Aug 2023 · 1 repository · arXiv:2308.16458Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
GPT has become financially literate: Insights from financial literacy tests of GPT and a preliminary test of how people use it as a source of advice 31 Aug 2023 · 0 repositories · arXiv:2309.00649