Methods › General › Learning Rate Schedules › Cosine Annealing › Papers, page 36
Cosine Annealing
Papers archive 2025-07-28
archive papers tagged: 3,965 · with a code link: 1,734 · where Syntology ran a sample: 627 (513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (627 of 3,965 tagged: 513 with a run with no instrument failure, 114 where every run was a failure of Syntology's instrument)
Page 36 of 40: papers 3,501 to 3,600 of 3,965, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
TruthfulQA: Measuring How Models Mimic Human Falsehoods 8 Sep 2021 · 3 repositories · arXiv:2109.07958Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Empathetic Dialogue Generation with Pre-trained RoBERTa-GPT2 and External Knowledge 7 Sep 2021 · 0 repositories · arXiv:2109.03004
-
NumGPT: Improving Numeracy Ability of Generative Pre-trained Models 7 Sep 2021 · 0 repositories · arXiv:2109.03137
-
Text-Free Prosody-Aware Generative Spoken Language Modeling 7 Sep 2021 · 1 repository · arXiv:2109.03264
-
General-Purpose Question-Answering with Macaw 6 Sep 2021 · 2 repositories · arXiv:2109.02593
-
GPT-3 Models are Poor Few-Shot Learners in the Biomedical Domain 6 Sep 2021 · 1 repository · arXiv:2109.02555
-
Finetuned Language Models Are Zero-Shot Learners 3 Sep 2021 · 8 repositories · arXiv:2109.01652Syntology official: harvested for another paper · 0 ran · 1 unverified (of 1 harvested sample)
-
Revisiting 3D ResNets for Video Recognition 3 Sep 2021 · 5 repositories · arXiv:2109.01696
-
CodeT5: Identifier-aware Unified Pre-trained Encoder-Decoder Models for Code Understanding and Generation 2 Sep 2021 · 5 repositories · arXiv:2109.00859Syntology official: harvested, nothing ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
ConQX: Semantic Expansion of Spoken Queries for Intent Detection based on Conditioned Text Generation 2 Sep 2021 · 0 repositories · arXiv:2109.00729
-
Evaluating the Single-Shot MultiBox Detector and YOLO Deep Learning Models for the Detection of Tomatoes in a Greenhouse 2 Sep 2021 · 0 repositories · arXiv:2109.00810
-
So Cloze yet so Far: N400 Amplitude is Better Predicted by Distributional Information than Human Predictability Judgements 2 Sep 2021 · 0 repositories · arXiv:2109.01226
-
Fight Fire with Fire: Fine-tuning Hate Detectors using Large Samples of Generated Hate Speech 1 Sep 2021 · 0 repositories · arXiv:2109.00591
-
OptAGAN: Entropy-based finetuning on text VAE-GAN 1 Sep 2021 · 1 repository · arXiv:2109.00239
-
MiniF2F: a cross-system benchmark for formal Olympiad-level mathematics 31 Aug 2021 · 4 repositories · arXiv:2109.00110Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Task-Oriented Dialogue System as Natural Language Generation 31 Aug 2021 · 1 repository · arXiv:2108.13679Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 12 harvested samples)
-
On the Multilingual Capabilities of Very Large-Scale English Language Models 30 Aug 2021 · 2 repositories · arXiv:2108.13349
-
Want To Reduce Labeling Cost? GPT-3 Can Help 30 Aug 2021 · 1 repository · arXiv:2108.13487
-
HeadlineCause: A Dataset of News Headlines for Detecting Causalities 28 Aug 2021 · 1 repository · arXiv:2108.12626
-
CGEMs: A Metric Model for Automatic Code Generation using GPT-3 23 Aug 2021 · 0 repositories · arXiv:2108.10168
-
Table Caption Generation in Scholarly Documents Leveraging Pre-trained Language Models 18 Aug 2021 · 1 repository · arXiv:2108.08111
-
The Stability-Efficiency Dilemma: Investigating Sequence Length Warmup for Training GPT Models 13 Aug 2021 · 1 repository · arXiv:2108.06084
-
AMMUS : A Survey of Transformer-based Pretrained Models in Natural Language Processing 12 Aug 2021 · 1 repository · arXiv:2108.05542
-
PatrickStar: Parallel Training of Pre-trained Models via Chunk-based Memory Management 12 Aug 2021 · 1 repository · arXiv:2108.05818
-
Offensive Language and Hate Speech Detection with Deep Learning and Transfer Learning 6 Aug 2021 · 0 repositories · arXiv:2108.03305
-
RockGPT: Reconstructing three-dimensional digital rocks from single two-dimensional slice from the perspective of video generation 5 Aug 2021 · 0 repositories · arXiv:2108.03132
-
Random Offset Block Embedding Array (ROBE) for CriteoTB Benchmark MLPerf DLRM Model : 1000× Compression and 3.1× Faster Inference 4 Aug 2021 · 0 repositories · arXiv:2108.02191
-
Q-Pain: A Question Answering Dataset to Measure Social Bias in Pain Management 3 Aug 2021 · 0 repositories · arXiv:2108.01764
-
BERTAC: Enhancing Transformer-based Language Models with Adversarially Pretrained Convolutional Neural Networks 1 Aug 2021 · 1 repository
-
Best of Both Worlds: Making High Accuracy Non-incremental Transformer-based Disfluency Detection Incremental 1 Aug 2021 · 0 repositories
-
Developing a Compressed Object Detection Model based on YOLOv4 for Deployment on Embedded GPU Platform of Autonomous System 1 Aug 2021 · 0 repositories · arXiv:2108.00392
-
Employing Argumentation Knowledge Graphs for Neural Argument Generation 1 Aug 2021 · 1 repository
-
Explanations for CommonsenseQA: New Dataset and Models 1 Aug 2021 · 0 repositories
-
KuiLeiXi: a Chinese Open-Ended Text Adventure Game 1 Aug 2021 · 0 repositories
-
PRAL: A Tailored Pre-Training Model for Task-Oriented Dialog Generation 1 Aug 2021 · 0 repositories
-
TGEA: An Error-Annotated Dataset and Benchmark Tasks for TextGeneration from Pretrained Language Models 1 Aug 2021 · 0 repositories
-
Unleash GPT-2 Power for Event Detection 1 Aug 2021 · 0 repositories
-
Adapting GPT, GPT-2 and BERT Language Models for Speech Recognition 29 Jul 2021 · 0 repositories · arXiv:2108.07789
-
FNetAR: Mixing Tokens with Autoregressive Fourier Transforms 22 Jul 2021 · 1 repository · arXiv:2107.10932
-
Chimera: Efficiently Training Large-Scale Neural Networks with Bidirectional Pipelines 14 Jul 2021 · 1 repository · arXiv:2107.06925
-
Real Time Pear Fruit Detection and Counting Using YOLOv4 Models and Deep SORT 14 Jul 2021 · 3 repositories
-
Scalable Memory Protection in the PENGLAI Enclave 14 Jul 2021 · 1 repository
-
Real-Time Pothole Detection Using Deep Learning 13 Jul 2021 · 0 repositories · arXiv:2107.06356
-
Transformers with multi-modal features and post-fusion context for e-commerce session-based recommendation 11 Jul 2021 · 0 repositories · arXiv:2107.05124
-
Evaluating Large Language Models Trained on Code 7 Jul 2021 · 13 repositories · arXiv:2107.03374Syntology official (archive's flag): 2 ran · 26 ran (of which 0 constructed an object rather than computing a result; 24 with no instrument failure: 1 honoured, 0 violated, 23 with no contract checked; 2 where Syntology's instrument failed) · 13 unverified (of 39 harvested samples) · 4 pointer-only (licence)
-
Not Quite 'Ask a Librarian': AI on the Nature, Value, and Future of LIS 7 Jul 2021 · 0 repositories · arXiv:2107.05383
-
ERNIE 3.0: Large-scale Knowledge Enhanced Pre-training for Language Understanding and Generation 5 Jul 2021 · 2 repositories · arXiv:2107.02137
-
Is GPT-3 Text Indistinguishable from Human Text? Scarecrow: A Framework for Scrutinizing Machine Text 2 Jul 2021 · 0 repositories · arXiv:2107.01294
-
Achieving Real-Time Object Detection on MobileDevices with Neural Pruning Search 28 Jun 2021 · 0 repositories · arXiv:2106.14943
-
What's in a Measurement? Using GPT-3 on SemEval 2021 Task 8 -- MeasEval 28 Jun 2021 · 0 repositories · arXiv:2106.14720
-
SymbolicGPT: A Generative Transformer Model for Symbolic Regression 27 Jun 2021 · 2 repositories · arXiv:2106.14131Syntology community repositories only · 9 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 9 harvested samples)
-
Toward Less Hidden Cost of Code Completion with Acceptance and Ranking Models 26 Jun 2021 · 0 repositories · arXiv:2106.13928
-
Process for Adapting Language Models to Society (PALMS) with Values-Targeted Datasets 18 Jun 2021 · 0 repositories · arXiv:2106.10328
-
LoRA: Low-Rank Adaptation of Large Language Models 17 Jun 2021 · 74 repositories · arXiv:2106.09685Syntology community repositories only · 51 ran (of which 19 constructed an object rather than computing a result; 44 with no instrument failure: 1 honoured, 0 violated, 43 with no contract checked; 7 where Syntology's instrument failed) · 33 unverified (of 84 harvested samples) · 30 pointer-only (licence)
-
ASR Adaptation for E-commerce Chatbots using Cross-Utterance Context and Multi-Task Language Modeling 15 Jun 2021 · 0 repositories · arXiv:2106.09532
-
Textual Data Distributions: Kullback Leibler Textual Distributions Contrasts on GPT-2 Generated Texts, with Supervised, Unsupervised Learning on Vaccine & Market Topics & Sentiment 15 Jun 2021 · 0 repositories · arXiv:2107.02025
-
GPT3-to-plan: Extracting plans from text using GPT-3 14 Jun 2021 · 1 repository · arXiv:2106.07131
-
Pre-Trained Models: Past, Present and Future 14 Jun 2021 · 0 repositories · arXiv:2106.07139
-
Generate, Annotate, and Learn: NLP with Synthetic Text 11 Jun 2021 · 1 repository · arXiv:2106.06168Syntology official (archive's flag): 1 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Small Object Detection for Near Real-Time Egocentric Perception in a Manual Assembly Scenario 11 Jun 2021 · 0 repositories · arXiv:2106.06403
-
Programming Puzzles 10 Jun 2021 · 3 repositories · arXiv:2106.05784
-
Engines of Power: Electricity, AI, and General-Purpose Military Transformations 8 Jun 2021 · 0 repositories · arXiv:2106.04338
-
TIMEDIAL: Temporal Commonsense Reasoning in Dialog 8 Jun 2021 · 1 repository · arXiv:2106.04571
-
Do Syntactic Probes Probe Syntax? Experiments with Jabberwocky Probing 4 Jun 2021 · 0 repositories · arXiv:2106.02559
-
Auto-tagging of Short Conversational Sentences using Transformer Methods 3 Jun 2021 · 0 repositories · arXiv:2106.01735
-
Multiscale Domain Adaptive YOLO for Cross-Domain Object Detection 2 Jun 2021 · 1 repository · arXiv:2106.01483
-
GPT Perdetry Test: Generating new meanings for new words 1 Jun 2021 · 0 repositories
-
Towards a Comprehensive Understanding and Accurate Evaluation of Societal Biases in Pre-Trained Transformers 1 Jun 2021 · 0 repositories
-
Knowledge Inheritance for Pre-trained Language Models 28 May 2021 · 2 repositories · arXiv:2105.13880
-
Generative Adversarial Imitation Learning for Empathy-based AI 27 May 2021 · 0 repositories · arXiv:2105.13328
-
Data Curation and Quality Assurance for Machine Learning-based Cyber Intrusion Detection 20 May 2021 · 1 repository · arXiv:2105.10041
-
Methods for Detoxification of Texts for the Russian Language 19 May 2021 · 3 repositories · arXiv:2105.09052
-
Neural Predictive Text for Grammatical Error Prevention 16 May 2021 · 0 repositories
-
SLGPT: Using Transfer Learning to Directly Generate Simulink Model Files and Find Bugs in the Simulink Toolchain 16 May 2021 · 1 repository · arXiv:2105.07465
-
BERT Busters: Outlier Dimensions that Disrupt Transformers 14 May 2021 · 0 repositories · arXiv:2105.06990
-
RetGen: A Joint framework for Retrieval and Grounded Text Generation Modeling 14 May 2021 · 1 repository · arXiv:2105.06597
-
BERT is to NLP what AlexNet is to CV: Can Pre-Trained Language Models Identify Analogies? 11 May 2021 · 1 repository · arXiv:2105.04949
-
EL-Attention: Memory Efficient Lossless Attention for Generation 11 May 2021 · 1 repository · arXiv:2105.04779Syntology official (archive's flag): 6 ran · 6 ran (of which 2 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Active Terahertz Imaging Dataset for Concealed Object Detection 8 May 2021 · 1 repository · arXiv:2105.03677
-
e-ViL: A Dataset and Benchmark for Natural Language Explanations in Vision-Language Tasks 8 May 2021 · 2 repositories · arXiv:2105.03761Syntology official (archive's flag): 5 ran · 7 ran (of which 3 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 8 unverified (of 15 harvested samples) · 15 pointer-only (licence)
-
Facial Emotion Recognition: State of the Art Performance on FER2013 8 May 2021 · 2 repositories · arXiv:2105.03588
-
DExperts: Decoding-Time Controlled Text Generation with Experts and Anti-Experts 7 May 2021 · 1 repository · arXiv:2105.03023Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
A Robotic Approach towards Quantifying Epipelagic Bound Plastic Using Deep Visual Models 5 May 2021 · 1 repository · arXiv:2105.01882
-
One Model to Rule them All: Towards Zero-Shot Learning for Databases 3 May 2021 · 0 repositories · arXiv:2105.00642
-
Unreasonable Effectiveness of Rule-Based Heuristics in Solving Russian SuperGLUE Tasks 3 May 2021 · 0 repositories · arXiv:2105.01192
-
Mitigating Political Bias in Language Models Through Reinforced Calibration 30 Apr 2021 · 0 repositories · arXiv:2104.14795
-
Unsupervised data augmentation for object detection 30 Apr 2021 · 0 repositories · arXiv:2104.14965
-
Entailment as Few-Shot Learner 29 Apr 2021 · 3 repositories · arXiv:2104.14690Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Extractive and Abstractive Explanations for Fact-Checking and Evaluation of News 27 Apr 2021 · 0 repositories · arXiv:2104.12918
-
UoT-UWF-PartAI at SemEval-2021 Task 5: Self Attention Based Bi-GRU with Multi-Embedding Representation for Toxicity Highlighter 27 Apr 2021 · 0 repositories · arXiv:2104.13164
-
Accounting for Agreement Phenomena in Sentence Comprehension with Transformer Language Models: Effects of Similarity-based Interference on Surprisal and Attention 26 Apr 2021 · 0 repositories · arXiv:2104.12874
-
Easy and Efficient Transformer : Scalable Inference Solution For large NLP model 26 Apr 2021 · 1 repository · arXiv:2104.12470
-
PanGu-α: Large-scale Autoregressive Pretrained Chinese Language Models with Auto-parallel Computation 26 Apr 2021 · 5 repositories · arXiv:2104.12369
-
Adapting Long Context NLM for ASR Rescoring in Conversational Agents 21 Apr 2021 · 0 repositories · arXiv:2104.11070
-
Analyzing COVID-19 Tweets with Transformer-based Language Models 20 Apr 2021 · 0 repositories · arXiv:2104.10259
-
Efficient pre-training objectives for Transformers 20 Apr 2021 · 0 repositories · arXiv:2104.09694
-
A Token-level Reference-free Hallucination Detection Benchmark for Free-form Text Generation 18 Apr 2021 · 2 repositories · arXiv:2104.08704Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity 18 Apr 2021 · 2 repositories · arXiv:2104.08786Syntology 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 7 unverified (of 10 harvested samples)
-
GPT3Mix: Leveraging Large-scale Language Models for Text Augmentation 18 Apr 2021 · 1 repository · arXiv:2104.08826Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Cross-Task Generalization via Natural Language Crowdsourcing Instructions 18 Apr 2021 · 3 repositories · arXiv:2104.08773Syntology official (archive's flag): 4 ran · 5 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 4 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)