Methods › General › Learning Rate Schedules › Inverse Square Root Schedule › Papers, page 6
Inverse Square Root Schedule
Papers archive 2025-07-28
archive papers tagged: 702 · with a code link: 349 · where Syntology ran a sample: 97 (83 with a run with no instrument failure, 14 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (97 of 702 tagged: 83 with a run with no instrument failure, 14 where every run was a failure of Syntology's instrument)
Page 6 of 8: papers 501 to 600 of 702, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
CodeAttack: Code-Based Adversarial Attacks for Pre-trained Programming Language Models 31 May 2022 · 2 repositories · arXiv:2206.00052
-
E2S2: Encoding-Enhanced Sequence-to-Sequence Pretraining for Language Understanding and Generation 30 May 2022 · 1 repository · arXiv:2205.14912
-
Conditional set generation using Seq2seq models 25 May 2022 · 0 repositories · arXiv:2205.12485
-
RobustLR: Evaluating Robustness to Logical Perturbation in Deductive Reasoning 25 May 2022 · 1 repository · arXiv:2205.12598Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples)
-
TaCube: Pre-computing Data Cubes for Answering Numerical-Reasoning Questions over Tabular Data 25 May 2022 · 1 repository · arXiv:2205.12682
-
EdiT5: Semi-Autoregressive Text-Editing with T5 Warm-Start 24 May 2022 · 0 repositories · arXiv:2205.12209
-
FLUTE: Figurative Language Understanding through Textual Explanations 24 May 2022 · 1 repository · arXiv:2205.12404
-
GeoMLAMA: Geo-Diverse Commonsense Probing on Multilingual Pre-Trained Language Models 24 May 2022 · 1 repository · arXiv:2205.12247
-
Medical Scientific Table-to-Text Generation with Human-in-the-Loop under the Data Sparsity Constraint 24 May 2022 · 0 repositories · arXiv:2205.12368
-
Workflow Discovery from Dialogues in the Low Data Regime 24 May 2022 · 1 repository · arXiv:2205.11690
-
A Question-Answer Driven Approach to Reveal Affirmative Interpretations from Verbal Negations 23 May 2022 · 1 repository · arXiv:2205.11467
-
BanglaNLG and BanglaT5: Benchmarks and Resources for Evaluating Low-Resource Natural Language Generation in Bangla 23 May 2022 · 2 repositories · arXiv:2205.11081
-
Life after BERT: What do Other Muppets Understand about Language? 21 May 2022 · 1 repository · arXiv:2205.10696
-
Evaluation of Transfer Learning for Polish with a Text-to-Text Model 18 May 2022 · 0 repositories · arXiv:2205.08808
-
PASH at TREC 2021 Deep Learning Track: Generative Enhanced Model for Multi-stage Ranking 18 May 2022 · 0 repositories · arXiv:2205.11245
-
M6-Rec: Generative Pretrained Language Models are Open-Ended Recommender Systems 17 May 2022 · 0 repositories · arXiv:2205.08084
-
SKILL: Structured Knowledge Infusion for Large Language Models 17 May 2022 · 0 repositories · arXiv:2205.08184
-
Downstream Transformer Generation of Question-Answer Pairs with Preprocessing and Postprocessing Pipelines 15 May 2022 · 1 repository · arXiv:2205.07387
-
RASAT: Integrating Relational Structures into Pretrained Seq2Seq Model for Text-to-SQL 14 May 2022 · 1 repository · arXiv:2205.06983Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
UL2: Unifying Language Learning Paradigms 10 May 2022 · 2 repositories · arXiv:2205.05131Syntology community repositories only · 15 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 16 harvested samples)
-
Vector Representations of Idioms in Conversational Systems 7 May 2022 · 0 repositories · arXiv:2205.03666
-
Two New Datasets for Italian-Language Abstractive Text Summarization 29 Apr 2022 · 1 repository
-
Modern Baselines for SPARQL Semantic Parsing 27 Apr 2022 · 1 repository · arXiv:2204.12793
-
Towards Arabic Sentence Simplification via Classification and Generative Approaches 20 Apr 2022 · 0 repositories · arXiv:2204.09292
-
MASSIVE: A 1M-Example Multilingual Natural Language Understanding Dataset with 51 Typologically-Diverse Languages 18 Apr 2022 · 6 repositories · arXiv:2204.08582Syntology official (archive's flag): 1 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 1 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
WordAlchemy: A transformer-based Reverse Dictionary 16 Apr 2022 · 0 repositories · arXiv:2204.10181
-
ML_LTU at SemEval-2022 Task 4: T5 Towards Identifying Patronizing and Condescending Language 15 Apr 2022 · 0 repositories · arXiv:2204.07432
-
Building Markovian Generative Architectures over Pretrained LM Backbones for Efficient Task-Oriented Dialog Systems 13 Apr 2022 · 2 repositories · arXiv:2204.06452
-
Enhance Incomplete Utterance Restoration by Joint Learning Token Extraction and Text Generation 8 Apr 2022 · 1 repository · arXiv:2204.03958
-
MMTAfrica: Multilingual Machine Translation for African Languages 8 Apr 2022 · 1 repository · arXiv:2204.04306
-
ByT5 model for massively multilingual grapheme-to-phoneme conversion 6 Apr 2022 · 1 repository · arXiv:2204.03067
-
Example-based Hypernetworks for Out-of-Distribution Generalization 27 Mar 2022 · 1 repository · arXiv:2203.14276
-
Ensembling and Knowledge Distilling of Large Sequence Taggers for Grammatical Error Correction 24 Mar 2022 · 1 repository · arXiv:2203.13064
-
Recommendation as Language Processing (RLP): A Unified Pretrain, Personalized Prompt & Predict Paradigm (P5) 24 Mar 2022 · 2 repositories · arXiv:2203.13366
-
AraBART: a Pretrained Arabic Sequence-to-Sequence Model for Abstractive Summarization 21 Mar 2022 · 0 repositories · arXiv:2203.10945
-
DQ-BART: Efficient Sequence-to-Sequence Model via Joint Distillation and Quantization 21 Mar 2022 · 2 repositories · arXiv:2203.11239Syntology official: harvested, nothing ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Coloring the Blank Slate: Pre-training Imparts a Hierarchical Inductive Bias to Sequence-to-sequence Models 17 Mar 2022 · 1 repository · arXiv:2203.09397Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
IT5: Text-to-text Pretraining for Italian Language Understanding and Generation 7 Mar 2022 · 3 repositories · arXiv:2203.03759
-
Parameter-Efficient Mixture-of-Experts Architecture for Pre-trained Language Models 2 Mar 2022 · 2 repositories · arXiv:2203.01104Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
E-LANG: Energy-Based Joint Inferencing of Super and Swift Language Models 1 Mar 2022 · 0 repositories · arXiv:2203.00748
-
HyperPrompt: Prompt-based Task-Conditioning of Transformers 1 Mar 2022 · 0 repositories · arXiv:2203.00759
-
Mixture-of-Experts with Expert Choice Routing 18 Feb 2022 · 0 repositories · arXiv:2202.09368
-
Probing Pretrained Models of Source Code 16 Feb 2022 · 1 repository · arXiv:2202.08975
-
Predicting on the Edge: Identifying Where a Larger Model Does Better 15 Feb 2022 · 0 repositories · arXiv:2202.07652
-
A multi-task semi-supervised framework for Text2Graph & Graph2Text 12 Feb 2022 · 1 repository · arXiv:2202.06041
-
HaT5: Hate Language Identification using Text-to-Text Transfer Transformer 11 Feb 2022 · 0 repositories · arXiv:2202.05690
-
Logical Reasoning for Task Oriented Dialogue Systems 8 Feb 2022 · 0 repositories · arXiv:2202.04161
-
A Comparative Study on Language Models for Task-Oriented Dialogue Systems 21 Jan 2022 · 1 repository · arXiv:2201.08687
-
Cheating Automatic Short Answer Grading: On the Adversarial Usage of Adjectives and Adverbs 20 Jan 2022 · 1 repository · arXiv:2201.08318
-
GAP-Gen: Guided Automatic Python Code Generation 19 Jan 2022 · 1 repository · arXiv:2201.08810
-
AllWOZ: Towards Multilingual Task-Oriented Dialog Systems for All 16 Jan 2022 · 0 repositories
-
AraBART: a Pretrained Arabic Sequence-to-Sequence Model for Abstractive Summarization 16 Jan 2022 · 0 repositories
-
Are Pretrained Multilingual Models Equally Fair Across Languages? 16 Jan 2022 · 0 repositories
-
Auto-regressive Text Generation with Pre-Trained Language Models: An Empirical Study on Question-type Short Text Generation 16 Jan 2022 · 0 repositories
-
Efficient Zero-Shot Semantic Parsing with Paraphrasing from Pretrained Language Models 16 Jan 2022 · 0 repositories
-
Exploring Example Selection for Few-shot Text-to-SQL Semantic Parsing 16 Jan 2022 · 0 repositories
-
Exploring the Low-Resource Transfer-Learning with mT5 model 16 Jan 2022 · 0 repositories
-
Natural Language Deduction through Search over Statement Compositions 16 Jan 2022 · 0 repositories · arXiv:2201.06028
-
Re2G: Retrieve, Rerank, Generate 16 Jan 2022 · 0 repositories
-
UnifiedSKG: Unifying and Multi-Tasking Structured Knowledge Grounding with Text-to-Text Language Models 16 Jan 2022 · 1 repository · arXiv:2201.05966Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Comparison of biomedical relationship extraction methods and models for knowledge graph creation 5 Jan 2022 · 0 repositories · arXiv:2201.01647
-
The University of Texas at Dallas HLTRI's Participation in EPIC-QA: Searching for Entailed Questions Revealing Novel Answer Nuggets 28 Dec 2021 · 0 repositories · arXiv:2112.13946
-
Your Answer is Incorrect... Would you like to know why? Introducing a Bilingual Short Answer Feedback Dataset 17 Dec 2021 · 0 repositories
-
CrossSum: Beyond English-Centric Cross-Lingual Summarization for 1,500+ Language Pairs 16 Dec 2021 · 1 repository · arXiv:2112.08804
-
DREAM: Improving Situational QA by First Elaborating the Situation 16 Dec 2021 · 1 repository · arXiv:2112.08656Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
AllWOZ: Towards Multilingual Task-Oriented Dialog Systems for All 15 Dec 2021 · 0 repositories · arXiv:2112.08333
-
LongT5: Efficient Text-To-Text Transformer for Long Sequences 15 Dec 2021 · 4 repositories · arXiv:2112.07916Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Improving Compositional Generalization with Latent Structure and Data Augmentation 14 Dec 2021 · 2 repositories · arXiv:2112.07610
-
Do Data-based Curricula Work? 13 Dec 2021 · 0 repositories · arXiv:2112.06510
-
Towards Neural Functional Program Evaluation 9 Dec 2021 · 0 repositories · arXiv:2112.04630
-
Controlling Conditional Language Models without Catastrophic Forgetting 1 Dec 2021 · 2 repositories · arXiv:2112.00791Syntology official (archive's flag): 3 ran · 7 ran (of which 1 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 8 harvested samples) · 4 pointer-only (licence)
-
Searching for Efficient Transformers for Language Modeling 1 Dec 2021 · 0 repositories
-
Chemical Identification and Indexing in PubMed Articles via BERT and Text-to-Text Approaches 30 Nov 2021 · 0 repositories · arXiv:2111.15622
-
Text Mining Drug/Chemical-Protein Interactions using an Ensemble of BERT and T5 Based Models 30 Nov 2021 · 0 repositories · arXiv:2111.15617
-
ExT5: Towards Extreme Multi-Task Scaling for Transfer Learning 22 Nov 2021 · 4 repositories · arXiv:2111.10952Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples)
-
A Comparative Study on Transfer Learning and Distance Metrics in Semantic Clustering over the COVID-19 Tweets 16 Nov 2021 · 0 repositories · arXiv:2111.08658
-
Coloring the Blank Slate: Pre-training Imparts a Hierarchical Inductive Bias to Sequence-to-sequence Models 16 Nov 2021 · 0 repositories
-
CST5: Data augmentation for Code-Switched Semantic Parsing 16 Nov 2021 · 0 repositories
-
DAML-ST5: Low Resource Style Transfer via Domain Adaptive Meta Learning 16 Nov 2021 · 0 repositories
-
FRSUM: Towards Faithful Abstractive Summarization via Enhancing Factual Robustness 16 Nov 2021 · 0 repositories
-
Generation of News Articles from Tweets : An Experiment 16 Nov 2021 · 0 repositories
-
GLM: General Language Model Pretraining with Autoregressive Blank Infilling 16 Nov 2021 · 0 repositories
-
Improving Compositional Generalization with Self-Training for Data-to-Text Generation 16 Nov 2021 · 0 repositories
-
Life after BERT: What do Other Muppets Understand about Language? 16 Nov 2021 · 0 repositories
-
Neural Keyphrase Generation: Analysis and Evaluation 16 Nov 2021 · 0 repositories
-
The Power of Prompt Tuning for Low-Resource Semantic Parsing 16 Nov 2021 · 0 repositories
-
Calculating Question Similarity is Enough: A New Method for KBQA Tasks 15 Nov 2021 · 0 repositories · arXiv:2111.07658
-
Say What? Collaborative Pop Lyric Generation Using Multitask Transfer Learning 15 Nov 2021 · 0 repositories · arXiv:2111.07592
-
Automated question generation and question answering from Turkish texts 11 Nov 2021 · 1 repository · arXiv:2111.06476
-
Ask me in your own words: paraphrasing for multitask question answering 27 Oct 2021 · 1 repository
-
Fast Model Editing at Scale 21 Oct 2021 · 3 repositories · arXiv:2110.11309Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
EncT5: A Framework for Fine-tuning T5 as Non-autoregressive Models 16 Oct 2021 · 1 repository · arXiv:2110.08426
-
Evaluation of Transfer Learning for Polish with a text-to-text model 16 Oct 2021 · 0 repositories
-
Improving Compositional Generalization with Self-Training for Data-to-Text Generation 16 Oct 2021 · 1 repository · arXiv:2110.08467
-
Multi-Task End-to-End Training Improves Conversational Recommendation 16 Oct 2021 · 0 repositories
-
Sharpness-Aware Minimization Improves Language Model Generalization 16 Oct 2021 · 0 repositories · arXiv:2110.08529
-
The Power of Prompt Tuning for Low-Resource Semantic Parsing 16 Oct 2021 · 0 repositories · arXiv:2110.08525
-
LFPT5: A Unified Framework for Lifelong Few-shot Language Learning Based on Prompt Tuning of T5 14 Oct 2021 · 1 repository · arXiv:2110.07298Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; the one sample that ran constructed an object rather than computing a result (of 3 harvested samples)
-
SpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing 14 Oct 2021 · 6 repositories · arXiv:2110.07205Syntology community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
KG-FiD: Infusing Knowledge Graph in Fusion-in-Decoder for Open-Domain Question Answering 8 Oct 2021 · 0 repositories · arXiv:2110.04330