Methods › General › Learning Rate Schedules › Cosine Annealing › Papers, page 24
Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,965 · with a code link: 1,734 · where Syntology ran a sample: 627 (513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (627 of 3,965 tagged: 513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument)
Page 24 of 40: papers 2,301 to 2,400 of 3,965, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
I-WAS: a Data Augmentation Method with GPT-2 for Simile Detection 8 Aug 2023 · 0 repositories · arXiv:2308.04109
-
Fact-Checking Generative AI: Ontology-Driven Biological Graphs for Disease-Gene Link Verification 7 Aug 2023 · 0 repositories · arXiv:2308.03929
-
Exploring ChatGPT's Empathic Abilities 7 Aug 2023 · 1 repository · arXiv:2308.03527
-
CORAL: Expert-Curated medical Oncology Reports to Advance Language Model Inference 7 Aug 2023 · 1 repository · arXiv:2308.03853
-
KITLM: Domain-Specific Knowledge InTegration into Language Models for Question Answering 7 Aug 2023 · 1 repository · arXiv:2308.03638
-
RCMHA: Relative Convolutional Multi-Head Attention for Natural Language Modelling 7 Aug 2023 · 1 repository · arXiv:2308.03429
-
Topological Interpretations of GPT-3 7 Aug 2023 · 0 repositories · arXiv:2308.03565
-
GPTScan: Detecting Logic Vulnerabilities in Smart Contracts by Combining GPT with Program Analysis 7 Aug 2023 · 1 repository · arXiv:2308.03314
-
"Kurosawa": A Script Writer's Assistant 6 Aug 2023 · 0 repositories · arXiv:2308.03122
-
TARJAMAT: Evaluation of Bard and ChatGPT on Machine Translation of Ten Arabic Varieties 6 Aug 2023 · 0 repositories · arXiv:2308.03051
-
ChatGPT for GTFS: Benchmarking LLMs on GTFS Understanding and Retrieval 4 Aug 2023 · 1 repository · arXiv:2308.02618
-
Explaining Relation Classification Models with Semantic Extents 4 Aug 2023 · 2 repositories · arXiv:2308.02193
-
GEMRec: Towards Generative Model Recommendation 4 Aug 2023 · 1 repository · arXiv:2308.02205
-
Baby Llama: knowledge distillation from an ensemble of teachers trained on a small dataset with no performance penalty 3 Aug 2023 · 1 repository · arXiv:2308.02019
-
Baby's CoThought: Leveraging Large Language Models for Enhanced Reasoning in Compact Models 3 Aug 2023 · 1 repository · arXiv:2308.01684
-
ClassEval: A Manually-Crafted Benchmark for Evaluating LLMs on Class-level Code Generation 3 Aug 2023 · 2 repositories · arXiv:2308.01861
-
Does Correction Remain A Problem For Large Language Models? 3 Aug 2023 · 0 repositories · arXiv:2308.01776
-
Holy Grail 2.0: From Natural Language to Constraint Models 3 Aug 2023 · 0 repositories · arXiv:2308.01589
-
From Sparse to Soft Mixtures of Experts 2 Aug 2023 · 5 repositories · arXiv:2308.00951Syntology official (archive's flag): 3 ran · 29 ran (of which 5 constructed an object rather than computing a result; 18 with no instrument failure: 3 honoured, 4 violated, 11 with no contract checked; 11 where Syntology's instrument failed) · 4 unverified (of 33 harvested samples) · 8 pointer-only (licence)
-
Leveraging Few-Shot Data Augmentation and Waterfall Prompting for Response Generation 2 Aug 2023 · 0 repositories · arXiv:2308.01080
-
The Paradigm Shifts in Artificial Intelligence 2 Aug 2023 · 0 repositories · arXiv:2308.02558
-
ChatMOF: An Autonomous AI System for Predicting and Generating Metal-Organic Frameworks 1 Aug 2023 · 0 repositories · arXiv:2308.01423
-
Instructed to Bias: Instruction-Tuned Language Models Exhibit Emergent Cognitive Bias 1 Aug 2023 · 1 repository · arXiv:2308.00225Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Towards Effective Ancient Chinese Translation: Dataset, Model, and Evaluation 1 Aug 2023 · 1 repository · arXiv:2308.00240
-
Does fine-tuning GPT-3 with the OpenAI API leak personally-identifiable information? 31 Jul 2023 · 1 repository · arXiv:2307.16382
-
HAGRID: A Human-LLM Collaborative Dataset for Generative Information-Seeking with Attribution 31 Jul 2023 · 1 repository · arXiv:2307.16883
-
No that's not what I meant: Handling Third Position Repair in Conversational Question Answering 31 Jul 2023 · 1 repository · arXiv:2307.16689
-
Ontology engineering with Large Language Models 31 Jul 2023 · 0 repositories · arXiv:2307.16699
-
Evaluating ChatGPT and GPT-4 for Visual Programming 30 Jul 2023 · 0 repositories · arXiv:2308.02522
-
SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension 30 Jul 2023 · 3 repositories · arXiv:2307.16125
-
A Critical Review of Large Language Models: Sensitivity, Bias, and the Path Toward Specialized AI 28 Jul 2023 · 0 repositories · arXiv:2307.15425
-
Beyond Reality: The Pivotal Role of Generative AI in the Metaverse 28 Jul 2023 · 0 repositories · arXiv:2308.06272
-
Med-HALT: Medical Domain Hallucination Test for Large Language Models 28 Jul 2023 · 1 repository · arXiv:2307.15343
-
VeriGen: A Large Language Model for Verilog Code Generation 28 Jul 2023 · 0 repositories · arXiv:2308.00708
-
Evaluating Generative Models for Graph-to-Text Generation 27 Jul 2023 · 1 repository · arXiv:2307.14712
-
Metric-Based In-context Learning: A Case Study in Text Simplification 27 Jul 2023 · 1 repository · arXiv:2307.14632
-
New Interaction Paradigm for Complex EDA Software Leveraging GPT 27 Jul 2023 · 1 repository · arXiv:2307.14740Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
TextManiA: Enriching Visual Feature by Text-driven Manifold Augmentation 27 Jul 2023 · 0 repositories · arXiv:2307.14611
-
CliniDigest: A Case Study in Large Language Model Based Large-Scale Summarization of Clinical Trial Descriptions 26 Jul 2023 · 0 repositories · arXiv:2307.14522
-
How User Language Affects Conflict Fatality Estimates in ChatGPT 26 Jul 2023 · 0 repositories · arXiv:2308.00072
-
Mental-LLM: Leveraging Large Language Models for Mental Health Prediction via Online Text Data 26 Jul 2023 · 1 repository · arXiv:2307.14385
-
YOLOBench: Benchmarking Efficient Object Detectors on Embedded Systems 26 Jul 2023 · 0 repositories · arXiv:2307.13901
-
GPT-3 Models are Few-Shot Financial Reasoners 25 Jul 2023 · 0 repositories · arXiv:2307.13617
-
Is GPT a Computational Model of Emotion? Detailed Analysis 25 Jul 2023 · 0 repositories · arXiv:2307.13779
-
Predicting Code Coverage without Execution 25 Jul 2023 · 1 repository · arXiv:2307.13383
-
How Does Naming Affect LLMs on Code Analysis Tasks? 24 Jul 2023 · 0 repositories · arXiv:2307.12488
-
Gradient-Based Word Substitution for Obstinate Adversarial Examples Generation in Language Models 24 Jul 2023 · 0 repositories · arXiv:2307.12507
-
The potential of LLMs for coding with low-resource and domain-specific programming languages 24 Jul 2023 · 0 repositories · arXiv:2307.13018
-
HateModerate: Testing Hate Speech Detectors against Content Moderation Policies 23 Jul 2023 · 1 repository · arXiv:2307.12418
-
Validation of a Zero-Shot Learning Natural Language Processing Tool for Data Abstraction from Unstructured Healthcare Data 23 Jul 2023 · 1 repository · arXiv:2308.00107
-
AIGC Empowering Telecom Sector White Paper_chinese 21 Jul 2023 · 0 repositories · arXiv:2307.11449
-
GPT-4 Can't Reason 21 Jul 2023 · 0 repositories · arXiv:2308.03762
-
An In-Depth Evaluation of Federated Learning on Biomedical Natural Language Processing 20 Jul 2023 · 2 repositories · arXiv:2307.11254
-
Generative Language Models on Nucleotide Sequences of Human Genes 20 Jul 2023 · 1 repository · arXiv:2307.10634
-
IvyGPT: InteractiVe Chinese pathwaY language model in medical domain 20 Jul 2023 · 1 repository · arXiv:2307.10512
-
LLM Cognitive Judgements Differ From Human 20 Jul 2023 · 1 repository · arXiv:2307.11787Syntology official: harvested, nothing ran · 0 ran · 3 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Of Models and Tin Men: A Behavioural Economics Study of Principal-Agent Problems in AI Alignment using Large-Language Models 20 Jul 2023 · 2 repositories · arXiv:2307.11137
-
Controlling Equational Reasoning in Large Language Models with Prompt Interventions 19 Jul 2023 · 0 repositories · arXiv:2307.09998
-
How is ChatGPT's behavior changing over time? 18 Jul 2023 · 4 repositories · arXiv:2307.09009Syntology official (archive's flag): 1 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 5 pointer-only (licence)
-
Unveiling Gender Bias in Terms of Profession Across LLMs: Analyzing and Addressing Sociological Implications 18 Jul 2023 · 0 repositories · arXiv:2307.09162
-
A mixed policy to improve performance of language models on math problems 17 Jul 2023 · 1 repository · arXiv:2307.08767
-
A Study on the Performance of Generative Pre-trained Transformer (GPT) in Simulating Depressed Individuals on the Standardized Depressive Symptom Scale 17 Jul 2023 · 0 repositories · arXiv:2307.08576
-
ChatGPT is Good but Bing Chat is Better for Vietnamese Students 17 Jul 2023 · 0 repositories · arXiv:2307.08272
-
GEAR: Augmenting Language Models with Generalizable and Efficient Tool Resolution 17 Jul 2023 · 1 repository · arXiv:2307.08775Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Using an LLM to Help With Code Understanding 17 Jul 2023 · 0 repositories · arXiv:2307.08177
-
Legal Syllogism Prompting: Teaching Large Language Models for Legal Judgment Prediction 17 Jul 2023 · 1 repository · arXiv:2307.08321
-
SentimentGPT: Exploiting GPT for Advanced Sentiment Analysis and its Departure from Current Machine Learning 16 Jul 2023 · 1 repository · arXiv:2307.10234
-
Coupling Large Language Models with Logic Programming for Robust and General Reasoning from Text 15 Jul 2023 · 1 repository · arXiv:2307.07696Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Large Language Models as Superpositions of Cultural Perspectives 15 Jul 2023 · 0 repositories · arXiv:2307.07870
-
Leveraging Large Language Models to Generate Answer Set Programs 15 Jul 2023 · 1 repository · arXiv:2307.07699Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Fairness of ChatGPT and the Role Of Explainable-Guided Prompts 14 Jul 2023 · 1 repository · arXiv:2307.11761
-
MorphPiece : A Linguistic Tokenizer for Large Language Models 14 Jul 2023 · 0 repositories · arXiv:2307.07262
-
A Study on Differentiable Logic and LLMs for EPIC-KITCHENS-100 Unsupervised Domain Adaptation Challenge for Action Recognition 2023 13 Jul 2023 · 0 repositories · arXiv:2307.06569
-
Agreement Tracking for Multi-Issue Negotiation Dialogues 13 Jul 2023 · 0 repositories · arXiv:2307.06524
-
Negated Complementary Commonsense using Large Language Models 13 Jul 2023 · 1 repository · arXiv:2307.06794
-
Ashaar: Automatic Analysis and Generation of Arabic Poetry Using Deep Learning Approaches 12 Jul 2023 · 1 repository · arXiv:2307.06218
-
Distilling Large Language Models for Biomedical Knowledge Extraction: A Case Study on Adverse Drug Events 12 Jul 2023 · 0 repositories · arXiv:2307.06439
-
Argumentative Segmentation Enhancement for Legal Summarization 11 Jul 2023 · 0 repositories · arXiv:2307.05081
-
DNAGPT: A Generalized Pre-trained Tool for Versatile DNA Sequence Analysis Tasks 11 Jul 2023 · 0 repositories · arXiv:2307.05628
-
Large Language Models 11 Jul 2023 · 0 repositories · arXiv:2307.05782
-
Named entity recognition using GPT for identifying comparable companies 11 Jul 2023 · 0 repositories · arXiv:2307.07420
-
Unleashing the Emergent Cognitive Synergy in Large Language Models: A Task-Solving Agent through Multi-Persona Self-Collaboration 11 Jul 2023 · 2 repositories · arXiv:2307.05300
-
AmadeusGPT: a natural language interface for interactive animal behavioral analysis 10 Jul 2023 · 1 repository · arXiv:2307.04858
-
SimpleMTOD: A Simple Language Model for Multimodal Task-Oriented Dialogue with Symbolic Scene Representation 10 Jul 2023 · 0 repositories · arXiv:2307.04907
-
Assessing the efficacy of large language models in generating accurate teacher responses 9 Jul 2023 · 0 repositories · arXiv:2307.04274
-
A Stitch in Time Saves Nine: Detecting and Mitigating Hallucinations of LLMs by Validating Low-Confidence Generation 8 Jul 2023 · 0 repositories · arXiv:2307.03987
-
DWReCO at CheckThat! 2023: Enhancing Subjectivity Detection through Style-based Data Sampling 7 Jul 2023 · 1 repository · arXiv:2307.03550
-
Goal-Conditioned Predictive Coding for Offline Reinforcement Learning 7 Jul 2023 · 0 repositories · arXiv:2307.03406
-
How does AI chat change search behaviors? 7 Jul 2023 · 0 repositories · arXiv:2307.03826
-
RADAR: Robust AI-Text Detection via Adversarial Learning 7 Jul 2023 · 0 repositories · arXiv:2307.03838
-
Improving Retrieval-Augmented Large Language Models via Data Importance Learning 6 Jul 2023 · 1 repository · arXiv:2307.03027Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
Large Language Models Empowered Autonomous Edge AI for Connected Intelligence 6 Jul 2023 · 0 repositories · arXiv:2307.02779
-
Text Alignment Is An Efficient Unified Model for Massive NLP Tasks 6 Jul 2023 · 1 repository · arXiv:2307.02729
-
CAME: Confidence-guided Adaptive Memory Efficient Optimization 5 Jul 2023 · 2 repositories · arXiv:2307.02047Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples) · 1 pointer-only (licence)
-
Evaluating the Effectiveness of Large Language Models in Representing Textual Descriptions of Geometry and Spatial Relations 5 Jul 2023 · 0 repositories · arXiv:2307.03678
-
External Reasoning: Towards Multi-Large-Language-Models Interchangeable Assistance with Human Feedback 5 Jul 2023 · 1 repository · arXiv:2307.12057
-
Hoodwinked: Deception and Cooperation in a Text-Based Game for Language Models 5 Jul 2023 · 1 repository · arXiv:2308.01404
-
Multilingual Controllable Transformer-Based Lexical Simplification 5 Jul 2023 · 1 repository · arXiv:2307.02120
-
Open-Source LLMs for Text Annotation: A Practical Guide for Model Setting and Fine-Tuning 5 Jul 2023 · 0 repositories · arXiv:2307.02179
-
The FormAI Dataset: Generative AI in Software Security Through the Lens of Formal Verification 5 Jul 2023 · 0 repositories · arXiv:2307.02192