Methods › General › Learning Rate Schedules › Cosine Annealing › Papers, page 31
Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,965 · with a code link: 1,734 · where Syntology ran a sample: 627 (513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (627 of 3,965 tagged: 513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument)
Page 31 of 40: papers 3,001 to 3,100 of 3,965, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
LLM-Planner: Few-Shot Grounded Planning for Embodied Agents with Large Language Models 8 Dec 2022 · 1 repository · arXiv:2212.04088
-
NP4G : Network Programming for Generalization 8 Dec 2022 · 1 repository · arXiv:2212.11118
-
The Role of AI in Drug Discovery: Challenges, Opportunities, and Strategies 8 Dec 2022 · 0 repositories · arXiv:2212.08104
-
DeepSpeed Data Efficiency: Improving Deep Learning Model Quality and Training Efficiency via Efficient Data Sampling and Routing 7 Dec 2022 · 1 repository · arXiv:2212.03597
-
Adaptive Testing of Computer Vision Models 6 Dec 2022 · 1 repository · arXiv:2212.02774Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Counterfactual reasoning: Do language models need world knowledge for causal understanding? 6 Dec 2022 · 1 repository · arXiv:2212.03278
-
MobilePTX: Sparse Coding for Pneumothorax Detection Given Limited Training Examples 6 Dec 2022 · 0 repositories · arXiv:2212.03282
-
Modern French Poetry Generation with RoBERTa and GPT-2 6 Dec 2022 · 0 repositories · arXiv:2212.02911
-
Audio-Driven Co-Speech Gesture Video Generation 5 Dec 2022 · 0 repositories · arXiv:2212.02350
-
Automatic Generation of Factual News Headlines in Finnish 5 Dec 2022 · 0 repositories · arXiv:2212.02170
-
Exploring the Limits of Differentially Private Deep Learning with Group-wise Clipping 3 Dec 2022 · 0 repositories · arXiv:2212.01539
-
SumREN: Summarizing Reported Speech about Events in News 2 Dec 2022 · 1 repository · arXiv:2212.01146
-
a survey on GPT-3 1 Dec 2022 · 0 repositories · arXiv:2212.00857
-
Distilling Reasoning Capabilities into Smaller Language Models 1 Dec 2022 · 1 repository · arXiv:2212.00193
-
Quadapter: Adapter for GPT-2 Quantization 30 Nov 2022 · 0 repositories · arXiv:2211.16912
-
Outfit Generation and Recommendation -- An Experimental Study 29 Nov 2022 · 0 repositories · arXiv:2211.16353
-
Prompted Opinion Summarization with GPT-3.5 29 Nov 2022 · 1 repository · arXiv:2211.15914
-
GPT-Neo for commonsense reasoning -- a theoretical and practical lens 28 Nov 2022 · 1 repository · arXiv:2211.15593
-
Scientific and Creative Analogies in Pretrained Language Models 28 Nov 2022 · 2 repositories · arXiv:2211.15268
-
Understanding BLOOM: An empirical study on diverse NLP tasks 27 Nov 2022 · 0 repositories · arXiv:2211.14865
-
GPT-3-driven pedagogical agents for training children's curious question-asking skills 25 Nov 2022 · 0 repositories · arXiv:2211.14228
-
PromptTTS: Controllable Text-to-Speech with Text Descriptions 22 Nov 2022 · 1 repository · arXiv:2211.12171
-
Exploring the Efficacy of Pre-trained Checkpoints in Text-to-Music Generation Task 21 Nov 2022 · 2 repositories · arXiv:2211.11216
-
Language in a Bottle: Language Model Guided Concept Bottlenecks for Interpretable Image Classification 21 Nov 2022 · 2 repositories · arXiv:2211.11158
-
PointCLIP V2: Prompting CLIP and GPT for Powerful 3D Open-world Learning 21 Nov 2022 · 2 repositories · arXiv:2211.11682Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 4 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 4 pointer-only (licence)
-
Conceptor-Aided Debiasing of Large Language Models 20 Nov 2022 · 0 repositories · arXiv:2211.11087
-
A survey on knowledge-enhanced multimodal learning 19 Nov 2022 · 0 repositories · arXiv:2211.12328
-
Ignore Previous Prompt: Attack Techniques For Language Models 17 Nov 2022 · 1 repository · arXiv:2211.09527Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
Random-LTD: Random and Layerwise Token Dropping Brings Efficient Training for Large-scale Transformers 17 Nov 2022 · 1 repository · arXiv:2211.11586
-
UniSumm and SummZoo: Unified Model and Diverse Benchmark for Few-Shot Summarization 17 Nov 2022 · 1 repository · arXiv:2211.09783
-
Galactica: A Large Language Model for Science 16 Nov 2022 · 1 repository · arXiv:2211.09085Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
TSMind: Alibaba and Soochow University's Submission to the WMT22 Translation Suggestion Task 16 Nov 2022 · 0 repositories · arXiv:2211.08987
-
GLUE-X: Evaluating Natural Language Understanding Models from an Out-of-distribution Generalization Perspective 15 Nov 2022 · 1 repository · arXiv:2211.08073
-
PromptCap: Prompt-Guided Task-Aware Image Captioning 15 Nov 2022 · 1 repository · arXiv:2211.09699
-
RobBERT-2022: Updating a Dutch Language Model to Account for Evolving Language Use 15 Nov 2022 · 0 repositories · arXiv:2211.08192
-
Are Hard Examples also Harder to Explain? A Study with Human and Model-Generated Explanations 14 Nov 2022 · 1 repository · arXiv:2211.07517Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
UGIF: UI Grounded Instruction Following 14 Nov 2022 · 0 repositories · arXiv:2211.07615
-
Textual Data Augmentation for Patient Outcomes Prediction 13 Nov 2022 · 0 repositories · arXiv:2211.06778
-
Large Language Models Meet Harry Potter: A Bilingual Dataset for Aligning Dialogue Agents with Characters 13 Nov 2022 · 1 repository · arXiv:2211.06869
-
On Optimizing the Communication of Model Parallelism 10 Nov 2022 · 0 repositories · arXiv:2211.05322
-
Collateral facilitation in humans and language models 9 Nov 2022 · 1 repository · arXiv:2211.05198Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Active Example Selection for In-Context Learning 8 Nov 2022 · 1 repository · arXiv:2211.04486Syntology official (archive's flag): 10 ran · 10 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 6 where Syntology's instrument failed) · 6 unverified (of 16 harvested samples)
-
Towards Asteroid Detection in Microlensing Surveys with Deep Learning 4 Nov 2022 · 1 repository · arXiv:2211.02239
-
Using Large Pre-Trained Language Model to Assist FDA in Premarket Medical Device 3 Nov 2022 · 0 repositories · arXiv:2212.01217
-
Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small 1 Nov 2022 · 7 repositories · arXiv:2211.00593Syntology community repositories only · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples)
-
Text-Only Training for Image Captioning using Noise-Injected CLIP 1 Nov 2022 · 4 repositories · arXiv:2211.00575Syntology official (archive's flag): 2 ran · 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers 31 Oct 2022 · 17 repositories · arXiv:2210.17323Syntology official (archive's flag): 1 ran · 5 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 10 unverified (of 15 harvested samples) · 1 pointer-only (licence)
-
SSD-LM: Semi-autoregressive Simplex-based Diffusion Language Model for Text Generation and Modular Control 31 Oct 2022 · 2 repositories · arXiv:2210.17432
-
Towards Zero-Shot and Few-Shot Table Question Answering using GPT-3 31 Oct 2022 · 0 repositories · arXiv:2210.17284
-
Learning to Decompose: Hypothetical Question Decomposition Based on Comparable Texts 30 Oct 2022 · 0 repositories · arXiv:2210.16865
-
Probing for targeted syntactic knowledge through grammatical error detection 28 Oct 2022 · 1 repository · arXiv:2210.16228
-
ROMA: Run-Time Object Detection To Maximize Real-Time Accuracy 28 Oct 2022 · 0 repositories · arXiv:2210.16083
-
COCO-DR: Combating Distribution Shifts in Zero-Shot Dense Retrieval with Contrastive and Distributionally Robust Learning 27 Oct 2022 · 1 repository · arXiv:2210.15212Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples)
-
TRScore: A Novel GPT-based Readability Scorer for ASR Segmentation and Punctuation model evaluation and selection 27 Oct 2022 · 0 repositories · arXiv:2210.15104
-
Exploring Robustness of Prefix Tuning in Noisy Data: A Case Study in Financial Sentiment Analysis 26 Oct 2022 · 0 repositories · arXiv:2211.05584
-
IELM: An Open Information Extraction Benchmark for Pre-Trained Language Models 25 Oct 2022 · 0 repositories · arXiv:2210.14128
-
XRICL: Cross-lingual Retrieval-Augmented In-Context Learning for Cross-lingual Text-to-SQL Semantic Parsing 25 Oct 2022 · 0 repositories · arXiv:2210.13693
-
Inferring Past Human Actions in Homes with Abductive Reasoning 24 Oct 2022 · 1 repository · arXiv:2210.13984
-
Emergent World Representations: Exploring a Sequence Model Trained on a Synthetic Task 24 Oct 2022 · 4 repositories · arXiv:2210.13382Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Exploring Euphemism Detection in Few-Shot and Zero-Shot Settings 24 Oct 2022 · 1 repository · arXiv:2210.12926
-
Perfectly Secure Steganography Using Minimum Entropy Coupling 24 Oct 2022 · 2 repositories · arXiv:2210.14889Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples)
-
Leveraging Large Language Models for Multiple Choice Question Answering 22 Oct 2022 · 1 repository · arXiv:2210.12353Syntology official (archive's flag): 6 ran · 6 ran (of which 3 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples)
-
Meta-learning Pathologies from Radiology Reports using Variance Aware Prototypical Networks 22 Oct 2022 · 0 repositories · arXiv:2210.13979
-
A Causal Framework to Quantify the Robustness of Mathematical Reasoning with Language Models 21 Oct 2022 · 1 repository · arXiv:2210.12023Syntology official (archive's flag): 5 ran · 5 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Diffuser: Efficient Transformers with Multi-hop Attention Diffusion for Long Sequences 21 Oct 2022 · 1 repository · arXiv:2210.11794
-
WikiWhy: Answering and Explaining Cause-and-Effect Questions 21 Oct 2022 · 0 repositories · arXiv:2210.12152
-
3DALL-E: Integrating Text-to-Image AI in 3D Design Workflows 20 Oct 2022 · 0 repositories · arXiv:2210.11603
-
Composing Ensembles of Pre-trained Models via Iterative Consensus 20 Oct 2022 · 0 repositories · arXiv:2210.11522
-
General Image Descriptors for Open World Image Retrieval using ViT CLIP 20 Oct 2022 · 1 repository · arXiv:2210.11141Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
BioGPT: Generative Pre-trained Transformer for Biomedical Text Generation and Mining 19 Oct 2022 · 4 repositories · arXiv:2210.10341
-
Towards a neural architecture of language: Deep learning versus logistics of access in neural architectures for compositional processing 19 Oct 2022 · 0 repositories · arXiv:2210.10543
-
Systematicity in GPT-3's Interpretation of Novel English Noun Compounds 18 Oct 2022 · 0 repositories · arXiv:2210.09492
-
Team Flow at DRC2022: Pipeline System for Travel Destination Recommendation Task in Spoken Dialogue 18 Oct 2022 · 0 repositories · arXiv:2210.09518
-
Tiny-Attention Adapter: Contexts Are More Important Than the Number of Parameters 18 Oct 2022 · 0 repositories · arXiv:2211.01979
-
A Generative User Simulator with GPT-based Architecture and Goal State Tracking for Reinforced Multi-Domain Dialog Systems 17 Oct 2022 · 1 repository · arXiv:2210.08692Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples)
-
Prompting GPT-3 To Be Reliable 17 Oct 2022 · 1 repository · arXiv:2210.09150
-
NormSAGE: Multi-Lingual Multi-Cultural Norm Discovery from Conversations On-the-Fly 16 Oct 2022 · 1 repository · arXiv:2210.08604
-
Object-Attentional Untargeted Adversarial Attack 16 Oct 2022 · 0 repositories · arXiv:2210.08472
-
DyLoRA: Parameter Efficient Tuning of Pre-trained Models using Dynamic Search-Free Low-Rank Adaptation 14 Oct 2022 · 2 repositories · arXiv:2210.07558Syntology official: not harvested · 0 ran · 1 unverified (of 1 harvested sample)
-
Extracting Cultural Commonsense Knowledge at Scale 14 Oct 2022 · 2 repositories · arXiv:2210.07763
-
"John is 50 years old, can his son be 65?" Evaluating NLP Models' Understanding of Feasibility 14 Oct 2022 · 1 repository · arXiv:2210.07471
-
TestAug: A Framework for Augmenting Capability-based NLP Tests 14 Oct 2022 · 1 repository · arXiv:2210.08097
-
Saliency Map Verbalization: Comparing Feature Importance Representations from Model-free and Instruction-based Methods 13 Oct 2022 · 1 repository · arXiv:2210.07222
-
Explanations from Large Language Models Make Small Reasoners Better 13 Oct 2022 · 0 repositories · arXiv:2210.06726
-
Jointly Reinforced User Simulator and Task-oriented Dialog System with Simplified Generative Architecture 13 Oct 2022 · 0 repositories · arXiv:2210.06706
-
Language Models of Code are Few-Shot Commonsense Learners 13 Oct 2022 · 2 repositories · arXiv:2210.07128Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Large Language Models are few(1)-shot Table Reasoners 13 Oct 2022 · 1 repository · arXiv:2210.06710
-
Are Sample-Efficient NLP Models More Robust? 12 Oct 2022 · 0 repositories · arXiv:2210.06456
-
Foundation Transformers 12 Oct 2022 · 4 repositories · arXiv:2210.06423Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
Predictive Querying for Autoregressive Neural Sequence Models 12 Oct 2022 · 1 repository · arXiv:2210.06464Syntology official (archive's flag): 7 ran · 7 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 5 where Syntology's instrument failed) · 6 unverified (of 13 harvested samples)
-
SUMBot: Summarizing Context in Open-Domain Dialogue Systems 12 Oct 2022 · 0 repositories · arXiv:2210.06496
-
REV: Information-Theoretic Evaluation of Free-Text Rationales 10 Oct 2022 · 1 repository · arXiv:2210.04982Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
The Minimum Wage as an Anchor: Effects on Determinations of Fairness by Humans and AI 10 Oct 2022 · 0 repositories · arXiv:2210.10585
-
ASDOT: Any-Shot Data-to-Text Generation with Pretrained Language Models 9 Oct 2022 · 1 repository · arXiv:2210.04325
-
Controllable Dialogue Simulation with In-Context Learning 9 Oct 2022 · 1 repository · arXiv:2210.04185
-
Fine-Tuning Pre-trained Transformers into Decaying Fast Weights 9 Oct 2022 · 1 repository · arXiv:2210.04243Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
AlphaTuning: Quantization-Aware Parameter-Efficient Adaptation of Large-Scale Pre-Trained Language Models 8 Oct 2022 · 0 repositories · arXiv:2210.03858
-
Automatic Chain of Thought Prompting in Large Language Models 7 Oct 2022 · 5 repositories · arXiv:2210.03493Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
How Large Language Models are Transforming Machine-Paraphrased Plagiarism 7 Oct 2022 · 3 repositories · arXiv:2210.03568
-
In Search of a Robust Facial Expressions Recognition Model: A Large-Scale Visual Cross-Corpus Study 7 Oct 2022 · 1 repository