Methods › General › Learning Rate Schedules › Cosine Annealing › Papers, page 20
Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,965 · with a code link: 1,734 · where Syntology ran a sample: 627 (513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (627 of 3,965 tagged: 513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument)
Page 20 of 40: papers 1,901 to 2,000 of 3,965, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Fumbling in Babel: An Investigation into ChatGPT's Language Identification Ability 16 Nov 2023 · 0 repositories · arXiv:2311.09696
-
Generative AI for Hate Speech Detection: Evaluation and Findings 16 Nov 2023 · 0 repositories · arXiv:2311.09993
-
Human Still Wins over LLM: An Empirical Study of Active Learning on Domain-Specific Annotation Tasks 16 Nov 2023 · 0 repositories · arXiv:2311.09825
-
INTERVENOR: Prompting the Coding Ability of Large Language Models with the Interactive Chain of Repair 16 Nov 2023 · 1 repository · arXiv:2311.09868
-
FinanceMath: Knowledge-Intensive Math Reasoning in Finance Domains 16 Nov 2023 · 1 repository · arXiv:2311.09797Syntology official (archive's flag): 12 ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
On Retrieval Augmentation and the Limitations of Language Model Training 16 Nov 2023 · 0 repositories · arXiv:2311.09615
-
Predictive Minds: LLMs As Atypical Active Inference Agents 16 Nov 2023 · 0 repositories · arXiv:2311.10215
-
Reducing Privacy Risks in Online Self-Disclosures with Language Models 16 Nov 2023 · 0 repositories · arXiv:2311.09538
-
Can Large Language Models Follow Concept Annotation Guidelines? A Case Study on Scientific and Financial Domains 15 Nov 2023 · 1 repository · arXiv:2311.08704
-
Evaluating Gender Bias in the Translation of Gender-Neutral Languages into English 15 Nov 2023 · 0 repositories · arXiv:2311.08836
-
LOKE: Linked Open Knowledge Extraction for Automated Knowledge Graph Construction 15 Nov 2023 · 0 repositories · arXiv:2311.09366
-
ToolTalk: Evaluating Tool-Usage in a Conversational Setting 15 Nov 2023 · 1 repository · arXiv:2311.10775Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
"We Demand Justice!": Towards Social Context Grounding of Political Texts 15 Nov 2023 · 1 repository · arXiv:2311.09106
-
Unifying the Perspectives of NLP and Software Engineering: A Survey on Language Models for Code 14 Nov 2023 · 1 repository · arXiv:2311.07989
-
CPopQA: Ranking Cultural Concept Popularity by LLMs 14 Nov 2023 · 0 repositories · arXiv:2311.07897
-
Evaluating LLMs on Document-Based QA: Exact Answer Selection and Numerical Extraction using Cogtale dataset 14 Nov 2023 · 0 repositories · arXiv:2311.07878
-
Fair Abstractive Summarization of Diverse Perspectives 14 Nov 2023 · 1 repository · arXiv:2311.07884
-
Language Models are Better Bug Detector Through Code-Pair Classification 14 Nov 2023 · 1 repository · arXiv:2311.07957
-
Large Language Model-Driven Classroom Flipping: Empowering Student-Centric Peer Questioning with Flipped Interaction 14 Nov 2023 · 0 repositories · arXiv:2311.14708
-
Memory-efficient Stochastic methods for Memory-based Transformers 14 Nov 2023 · 1 repository · arXiv:2311.08123
-
Do large language models and humans have similar behaviors in causal inference with script knowledge? 13 Nov 2023 · 1 repository · arXiv:2311.07311
-
In-context Learning Generalizes, But Not Always Robustly: The Case of Syntax 13 Nov 2023 · 1 repository · arXiv:2311.07811
-
It's Not Easy Being Wrong: Large Language Models Struggle with Process of Elimination Reasoning 13 Nov 2023 · 1 repository · arXiv:2311.07532
-
Language Model-In-The-Loop: Data Optimal Approach to Learn-To-Recommend Actions in Text Games 13 Nov 2023 · 0 repositories · arXiv:2311.07687
-
MEGAVERSE: Benchmarking Large Language Models Across Languages, Modalities, Models and Tasks 13 Nov 2023 · 0 repositories · arXiv:2311.07463
-
On The Truthfulness of 'Surprisingly Likely' Responses of Large Language Models 13 Nov 2023 · 0 repositories · arXiv:2311.07692
-
Speech-based Slot Filling using Large Language Models 13 Nov 2023 · 0 repositories · arXiv:2311.07418
-
STEER: Unified Style Transfer with Expert Reinforcement 13 Nov 2023 · 1 repository · arXiv:2311.07167
-
From Complex to Simple: Unraveling the Cognitive Tree for Reasoning with Small Language Models 12 Nov 2023 · 0 repositories · arXiv:2311.06754
-
GIELLM: Japanese General Information Extraction Large Language Model Utilizing Mutual Reinforcement Effect 12 Nov 2023 · 0 repositories · arXiv:2311.06838
-
Large Language Models are In-context Teachers for Knowledge Reasoning 12 Nov 2023 · 0 repositories · arXiv:2311.06985
-
Data Contamination Quiz: A Tool to Detect and Estimate Contamination in Large Language Models 10 Nov 2023 · 2 repositories · arXiv:2311.06233Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Establishing Performance Baselines in Fine-Tuning, Retrieval-Augmented Generation and Soft-Prompting for Non-Specialist LLM Users 10 Nov 2023 · 0 repositories · arXiv:2311.05903
-
Exploring Fine-tuning ChatGPT for News Recommendation 10 Nov 2023 · 0 repositories · arXiv:2311.05850
-
Smart Agent-Based Modeling: On the Use of Large Language Models in Computer Simulations 10 Nov 2023 · 4 repositories · arXiv:2311.06330Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
YOLOv5s-BC: An improved YOLOv5s-based method for real-time apple detection 10 Nov 2023 · 0 repositories · arXiv:2311.05811
-
GeoFormer: Predicting Human Mobility using Generative Pre-trained Transformer (GPT) 9 Nov 2023 · 0 repositories · arXiv:2311.05092
-
Large Language Models and Prompt Engineering for Biomedical Query Focused Multi-Document Summarisation 9 Nov 2023 · 0 repositories · arXiv:2311.05169
-
Leveraging Artificial Intelligence Technology for Mapping Research to Sustainable Development Goals: A Case Study 9 Nov 2023 · 0 repositories · arXiv:2311.16162
-
Agent Lumos: Unified and Modular Training for Open-Source Language Agents 9 Nov 2023 · 2 repositories · arXiv:2311.05657Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Vision Encoder-Decoder Models for AI Coaching 9 Nov 2023 · 2 repositories · arXiv:2311.16161
-
Massive Editing for Large Language Models via Meta Learning 8 Nov 2023 · 1 repository · arXiv:2311.04661Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Rethinking Benchmark and Contamination for Language Models with Rephrased Samples 8 Nov 2023 · 1 repository · arXiv:2311.04850Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Evaluating Large Language Models in Ophthalmology 7 Nov 2023 · 0 repositories · arXiv:2311.04933
-
Identifying and Mitigating Vulnerabilities in LLM-Integrated Applications 7 Nov 2023 · 0 repositories · arXiv:2311.16153
-
Towards Interpretable Sequence Continuation: Analyzing Shared Circuits in Large Language Models 7 Nov 2023 · 1 repository · arXiv:2311.04131Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Neuro-GPT: Towards A Foundation Model for EEG 7 Nov 2023 · 1 repository · arXiv:2311.03764Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
DeepInception: Hypnotize Large Language Model to Be Jailbreaker 6 Nov 2023 · 1 repository · arXiv:2311.03191Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
In-Context Learning for Knowledge Base Question Answering for Unmanned Systems based on Large Language Models 6 Nov 2023 · 0 repositories · arXiv:2311.02956
-
Unraveling Downstream Gender Bias from Large Language Models: A Study on AI Educational Writing Assistance 6 Nov 2023 · 1 repository · arXiv:2311.03311
-
Evaluating the Potential of Leading Large Language Models in Reasoning Biology Questions 5 Nov 2023 · 0 repositories · arXiv:2311.07582
-
Extraction of Atypical Aspects from Customer Reviews: Datasets and Experiments with Language Models 5 Nov 2023 · 2 repositories · arXiv:2311.02702
-
UID as a Guiding Metric for Automated Authorship Obfuscation 5 Nov 2023 · 0 repositories · arXiv:2312.03709
-
Automating Governing Knowledge Commons and Contextual Integrity (GKC-CI) Privacy Policy Annotations with Large Language Models 3 Nov 2023 · 1 repository · arXiv:2311.02192
-
COSMIC: Data Efficient Instruction-tuning For Speech In-Context Learning 3 Nov 2023 · 0 repositories · arXiv:2311.02248
-
Efficient Black-Box Adversarial Attacks on Neural Text Detectors 3 Nov 2023 · 1 repository · arXiv:2311.01873
-
Exploring the Numerical Reasoning Capabilities of Language Models: A Comprehensive Analysis on Tabular Data 3 Nov 2023 · 0 repositories · arXiv:2311.02216
-
Long Story Short: a Summarize-then-Search Method for Long Video Question Answering 2 Nov 2023 · 1 repository · arXiv:2311.01233
-
Measuring Five Accountable Talk Moves to Improve Instruction at Scale 2 Nov 2023 · 0 repositories · arXiv:2311.10749
-
Server-side Rescoring of Spoken Entity-centric Knowledge Queries for Virtual Assistants 2 Nov 2023 · 0 repositories · arXiv:2311.01398
-
Are Large Language Models Reliable Judges? A Study on the Factuality Evaluation Capabilities of LLMs 1 Nov 2023 · 0 repositories · arXiv:2311.00681
-
Continuous Training and Fine-tuning for Domain-Specific Language Models in Medical Question Answering 1 Nov 2023 · 0 repositories · arXiv:2311.00204
-
Is GPT Powerful Enough to Analyze the Emotions of Memes? 1 Nov 2023 · 0 repositories · arXiv:2311.00223
-
Unsupervised Lexical Simplification with Context Augmentation 1 Nov 2023 · 1 repository · arXiv:2311.00310
-
Do large language models solve verbal analogies like children do? 31 Oct 2023 · 0 repositories · arXiv:2310.20384
-
Does GPT-4 pass the Turing test? 31 Oct 2023 · 0 repositories · arXiv:2310.20216
-
Efficient Classification of Student Help Requests in Programming Courses Using Large Language Models 31 Oct 2023 · 0 repositories · arXiv:2310.20105
-
Interactive Multi-fidelity Learning for Cost-effective Adaptation of Language Model with Sparse Human Supervision 31 Oct 2023 · 0 repositories · arXiv:2310.20153
-
PsyCoT: Psychological Questionnaire as Powerful Chain-of-Thought for Personality Detection 31 Oct 2023 · 1 repository · arXiv:2310.20256
-
Theory of Mind in Large Language Models: Examining Performance of 11 State-of-the-Art models vs. Children Aged 7-10 on Advanced Tests 31 Oct 2023 · 0 repositories · arXiv:2310.20320
-
Herd: Using multiple, smaller LLMs to match the performances of proprietary, large LLMs via an intelligent composer 30 Oct 2023 · 0 repositories · arXiv:2310.19902
-
Interpretable-by-Design Text Understanding with Iteratively Generated Concept Bottleneck 30 Oct 2023 · 1 repository · arXiv:2310.19660
-
LitCab: Lightweight Language Model Calibration over Short- and Long-form Responses 30 Oct 2023 · 1 repository · arXiv:2310.19208
-
Remember what you did so you know what to do next 30 Oct 2023 · 0 repositories · arXiv:2311.01468
-
Synthetic Imitation Edit Feedback for Factual Alignment in Clinical Summarization 30 Oct 2023 · 1 repository · arXiv:2310.20033Syntology official (archive's flag): 7 ran · 7 ran (of which 3 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
EtiCor: Corpus for Analyzing LLMs for Etiquettes 29 Oct 2023 · 1 repository · arXiv:2310.18974
-
From Chatbots to PhishBots? -- Preventing Phishing scams created using ChatGPT, Google Bard and Claude 29 Oct 2023 · 0 repositories · arXiv:2310.19181
-
Efficient kernel surrogates for neural network-based regression 28 Oct 2023 · 0 repositories · arXiv:2310.18612
-
The Synergy of Speculative Decoding and Batching in Serving Large Language Models 28 Oct 2023 · 0 repositories · arXiv:2310.18813
-
Large language models for aspect-based sentiment analysis 27 Oct 2023 · 1 repository · arXiv:2310.18025Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Lost in Translation, Found in Spans: Identifying Claims in Multilingual Social Media 27 Oct 2023 · 1 repository · arXiv:2310.18205
-
OffMix-3L: A Novel Code-Mixed Dataset in Bangla-English-Hindi for Offensive Language Identification 27 Oct 2023 · 1 repository · arXiv:2310.18387
-
SentMix-3L: A Bangla-English-Hindi Code-Mixed Dataset for Sentiment Analysis 27 Oct 2023 · 1 repository · arXiv:2310.18023
-
Can LLMs Grade Short-Answer Reading Comprehension Questions : An Empirical Study with a Novel Dataset 26 Oct 2023 · 0 repositories · arXiv:2310.18373
-
FedPEAT: Convergence of Federated Learning, Parameter-Efficient Fine Tuning, and Emulator Assisted Tuning for Artificial Intelligence Foundation Models with Mobile Edge Computing 26 Oct 2023 · 0 repositories · arXiv:2310.17491
-
From Transcripts to Insights: Uncovering Corporate Risks Using Generative AI 26 Oct 2023 · 0 repositories · arXiv:2310.17721
-
In-Context Learning Dynamics with Random Binary Sequences 26 Oct 2023 · 1 repository · arXiv:2310.17639Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
LightLM: A Lightweight Deep and Narrow Language Model for Generative Recommendation 26 Oct 2023 · 1 repository · arXiv:2310.17488Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 1 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
DecoderTracker: Decoder-Only Method for Multiple-Object Tracking 26 Oct 2023 · 0 repositories · arXiv:2310.17170
-
Sliceformer: Make Multi-head Attention as Simple as Sorting in Discriminative Tasks 26 Oct 2023 · 1 repository · arXiv:2310.17683
-
"You Are An Expert Linguistic Annotator": Limits of LLMs as Analyzers of Abstract Meaning Representation 26 Oct 2023 · 0 repositories · arXiv:2310.17793
-
ZeroQuant-HERO: Hardware-Enhanced Robust Optimized Post-Training Quantization Framework for W8A8 Transformers 26 Oct 2023 · 0 repositories · arXiv:2310.17723
-
BabyStories: Can Reinforcement Learning Teach Baby Language Models to Write Better Stories? 25 Oct 2023 · 1 repository · arXiv:2310.16681Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
BOOST: Harnessing Black-Box Control to Boost Commonsense in LMs' Generation 25 Oct 2023 · 0 repositories · arXiv:2310.17054
-
Can GPT models Follow Human Summarization Guidelines? Evaluating ChatGPT and GPT-4 for Dialogue Summarization 25 Oct 2023 · 0 repositories · arXiv:2310.16810
-
Decoding Stumpers: Large Language Models vs. Human Problem-Solvers 25 Oct 2023 · 0 repositories · arXiv:2310.16411
-
Discrete Diffusion Modeling by Estimating the Ratios of the Data Distribution 25 Oct 2023 · 4 repositories · arXiv:2310.16834Syntology official (archive's flag): 16 ran · 16 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 1 honoured, 0 violated, 13 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 18 harvested samples) · 15 pointer-only (licence)
-
How well can machine-generated texts be identified and can language models be trained to avoid identification? 25 Oct 2023 · 0 repositories · arXiv:2310.16992
-
Muslim-Violence Bias Persists in Debiased GPT Models 25 Oct 2023 · 0 repositories · arXiv:2310.18368
-
R³ Prompting: Review, Rephrase and Resolve for Chain-of-Thought Reasoning in Large Language Models under Noisy Context 25 Oct 2023 · 0 repositories · arXiv:2310.16535