Methods › General › Learning Rate Schedules › Inverse Square Root Schedule › Papers, page 2
Inverse Square Root Schedule
Papers archive 2025-07-28
archive papers tagged: 702 · with a code link: 349 · where Syntology ran a sample: 97 (83 with a run with no instrument failure, 14 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (97 of 702 tagged: 83 with a run with no instrument failure, 14 where every run was a failure of Syntology's instrument)
Page 2 of 8: papers 101 to 200 of 702, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Masking The Bias : From Echo Chambers to Large Scale Aspect-Based Sentiment Analysis 2 Sep 2024 · 1 repository
-
MaFeRw: Query Rewriting with Multi-Aspect Feedbacks for Retrieval-Augmented Large Language Models 30 Aug 2024 · 1 repository · arXiv:2408.17072
-
Unintentional Security Flaws in Code: Automated Defense via Root Cause Analysis 30 Aug 2024 · 0 repositories · arXiv:2409.00199
-
MambaPlace:Text-to-Point-Cloud Cross-Modal Place Recognition with Attention Mamba Mechanisms 28 Aug 2024 · 1 repository · arXiv:2408.15740
-
Large Language Models as Foundations for Next-Gen Dense Retrieval: A Comprehensive Empirical Assessment 22 Aug 2024 · 0 repositories · arXiv:2408.12194
-
Factorized-Dreamer: Training A High-Quality Video Generator with Limited and Low-Quality Data 19 Aug 2024 · 0 repositories · arXiv:2408.10119
-
SSDTrain: An Activation Offloading Framework to SSDs for Faster Large Language Model Training 19 Aug 2024 · 0 repositories · arXiv:2408.10013
-
DataVisT5: A Pre-trained Language Model for Jointly Understanding Text and Data Visualization 14 Aug 2024 · 1 repository · arXiv:2408.07401
-
VulCatch: Enhancing Binary Vulnerability Detection through CodeT5 Decompilation and KAN Advanced Feature Extraction 13 Aug 2024 · 0 repositories · arXiv:2408.07181
-
Empathy Level Alignment via Reinforcement Learning for Empathetic Response Generation 6 Aug 2024 · 1 repository · arXiv:2408.02976
-
Automatic Pull Request Description Generation Using LLMs: A T5 Model Approach 1 Aug 2024 · 0 repositories · arXiv:2408.00921
-
Comparison of Large Language Models for Generating Contextually Relevant Questions 30 Jul 2024 · 1 repository · arXiv:2407.20578
-
Sentiment Analysis of Lithuanian Online Reviews Using Large Language Models 29 Jul 2024 · 0 repositories · arXiv:2407.19914
-
Positive Text Reframing under Multi-strategy Optimization 25 Jul 2024 · 1 repository · arXiv:2407.17940
-
Reporting and Analysing the Environmental Impact of Language Models on the Example of Commonsense Question Answering with External Knowledge 24 Jul 2024 · 0 repositories · arXiv:2408.01453
-
Promises and Pitfalls of Generative Masked Language Modeling: Theoretical Framework and Practical Guidelines 22 Jul 2024 · 1 repository · arXiv:2407.21046
-
Prior Knowledge Integration via LLM Encoding and Pseudo Event Regulation for Video Moment Retrieval 21 Jul 2024 · 1 repository · arXiv:2407.15051Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 1 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 11 harvested samples) · 9 pointer-only (licence)
-
On Initializing Transformers with Pre-trained Embeddings 17 Jul 2024 · 0 repositories · arXiv:2407.12514
-
LaMI-DETR: Open-Vocabulary Detection with Language Model Instruction 16 Jul 2024 · 1 repository · arXiv:2407.11335Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Weighted Grouped Query Attention in Transformers 15 Jul 2024 · 0 repositories · arXiv:2407.10855
-
YouTube-SL-25: A Large-Scale, Open-Domain Multilingual Sign Language Parallel Corpus 15 Jul 2024 · 1 repository · arXiv:2407.11144
-
Enhancing Emotion Prediction in News Headlines: Insights from ChatGPT and Seq2Seq Models for Free-Text Generation 14 Jul 2024 · 0 repositories · arXiv:2407.10091
-
Surgical Text-to-Image Generation 12 Jul 2024 · 0 repositories · arXiv:2407.09230
-
Segment-Based Interactive Machine Translation for Pre-trained Models 9 Jul 2024 · 0 repositories · arXiv:2407.06990
-
Vision-Braille: An End-to-End Tool for Chinese Braille Image-to-Text Translation 8 Jul 2024 · 0 repositories · arXiv:2407.06048
-
RDBE: Reasoning Distillation-Based Evaluation Enhances Automatic Essay Scoring 3 Jul 2024 · 0 repositories · arXiv:2407.13781
-
The Model Arena for Cross-lingual Sentiment Analysis: A Comparative Study in the Era of Large Language Models 27 Jun 2024 · 0 repositories · arXiv:2406.19358
-
Relation Extraction with Fine-Tuned Large Language Models in Retrieval Augmented Generation Frameworks 20 Jun 2024 · 0 repositories · arXiv:2406.14745
-
ptt5-v2: A Closer Look at Continued Pretraining of T5 Models for the Portuguese Language 16 Jun 2024 · 0 repositories · arXiv:2406.10806
-
RE-RAG: Improving Open-Domain QA Performance and Interpretability with Relevance Estimator in Retrieval-Augmented Generation 9 Jun 2024 · 1 repository · arXiv:2406.05794
-
Concept Formation and Alignment in Language Models: Bridging Statistical Patterns in Latent Space to Concept Taxonomy 8 Jun 2024 · 0 repositories · arXiv:2406.05315
-
DiNeR: a Large Realistic Dataset for Evaluating Compositional Generalization 7 Jun 2024 · 1 repository · arXiv:2406.04669
-
Heidelberg-Boston @ SIGTYP 2024 Shared Task: Enhancing Low-Resource Language Analysis With Character-Aware Hierarchical Transformers 30 May 2024 · 1 repository · arXiv:2405.20145
-
KerasCV and KerasNLP: Vision and Language Power-Ups 30 May 2024 · 0 repositories · arXiv:2405.20247
-
Faster Cascades via Speculative Decoding 29 May 2024 · 0 repositories · arXiv:2405.19261
-
Offline Regularised Reinforcement Learning for Large Language Models Alignment 29 May 2024 · 0 repositories · arXiv:2405.19107
-
DeeperImpact: Optimizing Sparse Learned Index Structures 27 May 2024 · 1 repository · arXiv:2405.17093
-
Parameter-free Clipped Gradient Descent Meets Polyak 23 May 2024 · 0 repositories · arXiv:2405.15010
-
IGOT: Information Gain Optimized Tokenizer on Domain Adaptive Pretraining 16 May 2024 · 0 repositories · arXiv:2405.09857
-
DEPTH: Discourse Education through Pre-Training Hierarchically 13 May 2024 · 1 repository · arXiv:2405.07788
-
Evaluating Text Summaries Generated by Large Language Models Using OpenAI's GPT 7 May 2024 · 0 repositories · arXiv:2405.04053
-
Utilizing GPT to Enhance Text Summarization: A Strategy to Minimize Hallucinations 7 May 2024 · 0 repositories · arXiv:2405.04039
-
Assessing Adversarial Robustness of Large Language Models: An Empirical Study 4 May 2024 · 0 repositories · arXiv:2405.02764
-
UQA: Corpus for Urdu Question Answering 2 May 2024 · 3 repositories · arXiv:2405.01458
-
Opinion Mining Using Pre-Trained Large Language Models: Identifying the Type, Polarity, Intensity, Expression, and Source of Private States 1 May 2024 · 1 repository
-
TuBA: Cross-Lingual Transferability of Backdoor Attacks in LLMs with Instruction Tuning 30 Apr 2024 · 0 repositories · arXiv:2404.19597
-
IndicGenBench: A Multilingual Benchmark to Evaluate Generation Capabilities of LLMs on Indic Languages 25 Apr 2024 · 1 repository · arXiv:2404.16816
-
Retrieval-Augmented Generation-based Relation Extraction 20 Apr 2024 · 1 repository · arXiv:2404.13397
-
Augmenting emotion features in irony detection with Large language modeling 18 Apr 2024 · 0 repositories · arXiv:2404.12291
-
Comparative Analysis of Deep Natural Networks and Large Language Models for Aspect-Based Sentiment Analysis 17 Apr 2024 · 1 repository
-
HLTCOE at TREC 2023 NeuCLIR Track 11 Apr 2024 · 0 repositories · arXiv:2404.08118
-
Medical mT5: An Open-Source Multilingual Text-to-Text LLM for The Medical Domain 11 Apr 2024 · 0 repositories · arXiv:2404.07613
-
Control-DAG: Constrained Decoding for Non-Autoregressive Directed Acyclic T5 using Weighted Finite State Automata 10 Apr 2024 · 1 repository · arXiv:2404.06854
-
Data Bias According to Bipol: Men are Naturally Right and It is the Role of Women to Follow Their Lead 7 Apr 2024 · 1 repository · arXiv:2404.04838
-
Adaptive Cross-lingual Text Classification through In-Context One-Shot Demonstrations 3 Apr 2024 · 1 repository · arXiv:2404.02452Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
On Linearizing Structured Data in Encoder-Decoder Language Models: Insights from Text-to-SQL 3 Apr 2024 · 0 repositories · arXiv:2404.02389
-
Mitigating Misleading Chain-of-Thought Reasoning with Selective Filtering 28 Mar 2024 · 1 repository · arXiv:2403.19167
-
Reshaping Free-Text Radiology Notes Into Structured Reports With Generative Transformers 27 Mar 2024 · 1 repository · arXiv:2403.18938
-
Multilingual Sentence-T5: Scalable Sentence Encoders for Multilingual Applications 26 Mar 2024 · 0 repositories · arXiv:2403.17528
-
Transcribing Bengali Text with Regional Dialects to IPA using District Guided Tokens 26 Mar 2024 · 0 repositories · arXiv:2403.17407
-
Concurrent Linguistic Error Detection (CLED) for Large Language Models 25 Mar 2024 · 0 repositories · arXiv:2403.16393
-
R3CD: Scene Graph to Image Generation with Relation-aware Compositional Contrastive Control Diffusion 24 Mar 2024 · 0 repositories
-
SensoryT5: Infusing Sensorimotor Norms into T5 for Enhanced Fine-grained Emotion Classification 22 Mar 2024 · 0 repositories · arXiv:2403.15574
-
Automatic Summarization of Doctor-Patient Encounter Dialogues Using Large Language Model through Prompt Tuning 19 Mar 2024 · 0 repositories · arXiv:2403.13089
-
Evaluating Named Entity Recognition: A comparative analysis of mono- and multilingual transformer models on a novel Brazilian corporate earnings call transcripts dataset 18 Mar 2024 · 2 repositories · arXiv:2403.12212
-
Empirical Studies of Parameter Efficient Methods for Large Language Models of Code and Knowledge Transfer to R 16 Mar 2024 · 1 repository · arXiv:2405.01553
-
ExeGPT: Constraint-Aware Resource Scheduling for LLM Inference 15 Mar 2024 · 0 repositories · arXiv:2404.07947
-
Basque and Spanish Counter Narrative Generation: Data Creation and Evaluation 14 Mar 2024 · 0 repositories · arXiv:2403.09159
-
Information Extraction: An application to the domain of hyper-local financial data on developing countries 14 Mar 2024 · 0 repositories · arXiv:2403.09077
-
Autoregressive Score Generation for Multi-trait Essay Scoring 13 Mar 2024 · 1 repository · arXiv:2403.08332
-
Embedded Translations for Low-resource Automated Glossing 13 Mar 2024 · 0 repositories · arXiv:2403.08189
-
Chronos: Learning the Language of Time Series 12 Mar 2024 · 6 repositories · arXiv:2403.07815Syntology official (archive's flag): 8 ran · 23 ran (of which 0 constructed an object rather than computing a result; 22 with no instrument failure: 3 honoured, 1 violated, 18 with no contract checked; 1 where Syntology's instrument failed) · 5 unverified (of 28 harvested samples) · 5 pointer-only (licence)
-
Contextual Clarity: Generating Sentences with Transformer Models using Context-Reverso Data 12 Mar 2024 · 1 repository · arXiv:2403.08103
-
Linguistic Knowledge Can Enhance Encoder-Decoder Models (If You Let It) 27 Feb 2024 · 1 repository · arXiv:2402.17608
-
SKT5SciSumm -- Revisiting Extractive-Generative Approach for Multi-Document Scientific Summarization 27 Feb 2024 · 0 repositories · arXiv:2402.17311
-
ESG Sentiment Analysis: comparing human and language model performance including GPT 26 Feb 2024 · 0 repositories · arXiv:2402.16650
-
Advancing Parameter Efficiency in Fine-tuning via Representation Editing 23 Feb 2024 · 2 repositories · arXiv:2402.15179Syntology official (archive's flag): 1 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples) · 1 pointer-only (licence)
-
Annotation and Classification of Relevant Clauses in Terms-and-Conditions Contracts 22 Feb 2024 · 1 repository · arXiv:2402.14457
-
UMBCLU at SemEval-2024 Task 1A and 1C: Semantic Textual Relatedness with and without machine translation 20 Feb 2024 · 1 repository · arXiv:2402.12730
-
A synthetic data approach for domain generalization of NLI models 19 Feb 2024 · 0 repositories · arXiv:2402.12368
-
Key ingredients for effective zero-shot cross-lingual knowledge transfer in generative tasks 19 Feb 2024 · 0 repositories · arXiv:2402.12279
-
Emerging Opportunities of Using Large Language Models for Translation Between Drug Molecules and Indications 14 Feb 2024 · 0 repositories · arXiv:2402.09588
-
FGeo-TP: A Language Model-Enhanced Solver for Geometry Problems 14 Feb 2024 · 0 repositories · arXiv:2402.09047
-
Improving Black-box Robustness with In-Context Rewriting 13 Feb 2024 · 1 repository · arXiv:2402.08225
-
InkSight: Offline-to-Online Handwriting Conversion by Learning to Read and Write 8 Feb 2024 · 1 repository · arXiv:2402.05804
-
Lens: A Foundation Model for Network Traffic 6 Feb 2024 · 0 repositories · arXiv:2402.03646
-
CorpusLM: Towards a Unified Language Model on Corpus for Knowledge-Intensive Tasks 2 Feb 2024 · 0 repositories · arXiv:2402.01176
-
What Will My Model Forget? Forecasting Forgotten Examples in Language Model Refinement 2 Feb 2024 · 0 repositories · arXiv:2402.01865
-
Improving Semantic Control in Discrete Latent Spaces with Transformer Quantized Variational Autoencoders 1 Feb 2024 · 1 repository · arXiv:2402.00723
-
ToPro: Token-Level Prompt Decomposition for Cross-Lingual Sequence Labeling Tasks 29 Jan 2024 · 1 repository · arXiv:2401.16589
-
LPNL: Scalable Link Prediction with Large Language Models 24 Jan 2024 · 0 repositories · arXiv:2401.13227
-
APT: Adaptive Pruning and Tuning Pretrained Language Models for Efficient Training and Inference 22 Jan 2024 · 1 repository · arXiv:2401.12200Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
The Right Model for the Job: An Evaluation of Legal Multi-Label Classification Baselines 22 Jan 2024 · 0 repositories · arXiv:2401.11852
-
Finding a Needle in the Adversarial Haystack: A Targeted Paraphrasing Approach For Uncovering Edge Cases with Minimal Distribution Distortion 21 Jan 2024 · 1 repository · arXiv:2401.11373
-
LangBridge: Multilingual Reasoning Without Multilingual Supervision 19 Jan 2024 · 1 repository · arXiv:2401.10695Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 7 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Deciphering Textual Authenticity: A Generalized Strategy through the Lens of Large Language Semantics for Detecting Human vs. Machine-Generated Text 17 Jan 2024 · 1 repository · arXiv:2401.09407
-
Mapping Transformer Leveraged Embeddings for Cross-Lingual Document Representation 12 Jan 2024 · 1 repository · arXiv:2401.06583
-
PizzaCommonSense: Learning to Model Commonsense Reasoning about Intermediate Steps in Cooking Recipes 12 Jan 2024 · 1 repository · arXiv:2401.06930
-
AST-T5: Structure-Aware Pretraining for Code Generation and Understanding 5 Jan 2024 · 1 repository · arXiv:2401.03003Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
Are LLMs Robust for Spoken Dialogues? 4 Jan 2024 · 0 repositories · arXiv:2401.02297