Browse State-of-the-Art › Language Modeling › Papers, page 120
Language Modeling
Papers archive 2025-07-28
archive papers tagged: 14,182 · with a code link: 5,620 · where Syntology ran a sample: 1,894 (1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,894 of 14,182 tagged: 1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument)
Page 120 of 142: papers 11,901 to 12,000 of 14,182, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Learning to Selectively Learn for Weakly-supervised Paraphrase Generation25 Sep 2021 0 repositories listed
-
A Diversity-Enhanced and Constraints-Relaxed Augmentation for Low-Resource Classification24 Sep 2021 0 repositories listed
-
Identification of Enzymatic Active Sites with Unsupervised Language Modeling24 Sep 2021 0 repositories listed
-
MLIM: Vision-and-Language Model Pre-training with Masked Language and Image Modeling24 Sep 2021 0 repositories listed
-
Predicting Attention Sparsity in Transformers24 Sep 2021 0 repositories listed
-
Cross-Lingual Language Model Meta-Pretraining23 Sep 2021 0 repositories listed
-
LSTM Hyper-Parameter Selection for Malware Detection: Interaction Effects and Hierarchical Selection Approach23 Sep 2021 0 repositories listed
-
BFClass: A Backdoor-free Text Classification Framework22 Sep 2021 0 repositories listed
-
Low-Latency Incremental Text-to-Speech Synthesis with Distilled Context Prediction Network22 Sep 2021 0 repositories listed
-
BERTweetFR : Domain Adaptation of Pre-Trained Language Models for French Tweets21 Sep 2021 0 repositories listed
-
Learning Domain Specific Language Models for Automatic Speech Recognition through Machine Translation21 Sep 2021 0 repositories listed
-
The Trade-offs of Domain Adaptation for Neural Language Models21 Sep 2021 0 repositories listed
-
Influence of ASR and Language Model on Alzheimer's Disease Detection20 Sep 2021 0 repositories listed
-
Learning Natural Language Generation from Scratch20 Sep 2021 0 repositories listed
-
Adversarial Training with Contrastive Learning in NLP19 Sep 2021 0 repositories listed
-
Wav-BERT: Cooperative Acoustic and Linguistic Representation Learning for Low-Resource Speech Recognition19 Sep 2021 0 repositories listed
-
Multilingual Molecular Representation Learning via Contrastive Pre-training18 Sep 2021 0 repositories listed
-
BART-light: One Decoder Layer Is Enough17 Sep 2021 0 repositories listed
-
Commonsense Knowledge-Augmented Pretrained Language Models for Causal Reasoning Classification17 Sep 2021 0 repositories listed
-
Exploring Multitask Learning for Low-Resource AbstractiveSummarization17 Sep 2021 0 repositories listed
-
Long-Range Modeling of Source Code Files with eWASH: Extended Window Access by Syntax Hierarchy17 Sep 2021 0 repositories listed
-
Machine Reading Comprehension: Generative or Extractive Reader?17 Sep 2021 0 repositories listed
-
Relating Neural Text Degeneration to Exposure Bias17 Sep 2021 0 repositories listed
-
SentiPrompt: Sentiment Knowledge Enhanced Prompt-Tuning for Aspect-Based Sentiment Analysis17 Sep 2021 0 repositories listed
-
A Bag of Tricks for Dialogue Summarization16 Sep 2021 0 repositories listed
-
Do Language Models Know the Way to Rome?16 Sep 2021 0 repositories listed
-
Let the CAT out of the bag: Contrastive Attributed explanations for Text16 Sep 2021 0 repositories listed
-
Regularized Training of Nearest Neighbor Language Models16 Sep 2021 0 repositories listed
-
The Language Model Understood the Prompt was Ambiguous: Probing Syntactic Uncertainty Through Generation16 Sep 2021 0 repositories listed
-
Beyond Glass-Box Features: Uncertainty Quantification Enhanced Quality Estimation for Neural Machine Translation15 Sep 2021 0 repositories listed
-
Improving Text Auto-Completion with Next Phrase Prediction15 Sep 2021 0 repositories listed
-
On the Complementarity of Data Selection and Fine Tuning for Domain Adaptation15 Sep 2021 0 repositories listed
-
RankNAS: Efficient Neural Architecture Search by Pairwise Ranking15 Sep 2021 0 repositories listed
-
Tied & Reduced RNN-T Decoder15 Sep 2021 0 repositories listed
-
KroneckerBERT: Learning Kronecker Decomposition for Pre-trained Language Models via Knowledge Distillation13 Sep 2021 0 repositories listed
-
Single-Read Reconstruction for DNA Data Storage Using Transformers12 Sep 2021 0 repositories listed
-
Dual-State Capsule Networks for Text Classification10 Sep 2021 0 repositories listed
-
EfficientCLIP: Efficient Cross-Modal Pre-training by Ensemble Confident Learning and Language Modeling10 Sep 2021 0 repositories listed
-
MetaXT: Meta Cross-Task Transfer between Disparate Label Spaces9 Sep 2021 0 repositories listed
-
8 Sep 2021 0 repositories listed
-
Sustainable Modular Debiasing of Language Models8 Sep 2021 0 repositories listed
-
7 Sep 2021 0 repositories listed
-
Rare Tokens Degenerate All Tokens: Improving Neural Text Generation via Adaptive Gradient Gating for Rare Token Embeddings7 Sep 2021 0 repositories listed
-
You should evaluate your language model on marginal likelihood over tokenisations6 Sep 2021 0 repositories listed
-
Language Modeling, Lexical Translation, Reordering: The Training Process of NMT through the Lens of Classical SMT3 Sep 2021 0 repositories listed
-
No Need to Know Everything! Efficiently Augmenting Language Models With External Knowledge3 Sep 2021 0 repositories listed
-
An Empirical Exploration in Quality Filtering of Text Data2 Sep 2021 0 repositories listed
-
ConQX: Semantic Expansion of Spoken Queries for Intent Detection based on Conditioned Text Generation2 Sep 2021 0 repositories listed
-
LegaLMFiT: Efficient Short Legal Text Classification with LSTM Language Model Pre-Training2 Sep 2021 0 repositories listed
-
Multimodal Conditionality for Natural Language Generation2 Sep 2021 0 repositories listed
-
Pre-training Language Model Incorporating Domain-specific Heterogeneous Knowledge into A Unified Representation2 Sep 2021 0 repositories listed
-
Developing a Clinical Language Model for Swedish: Continued Pretraining of Generic BERT with In-Domain Data1 Sep 2021 0 repositories listed
-
Does Knowledge Help General NLU? An Empirical Study1 Sep 2021 0 repositories listed
-
Domain-Specific Japanese ELECTRA Model Using a Small Corpus1 Sep 2021 0 repositories listed
-
Improving Character-Aware Neural Language Model by Warming up Character Encoder under Skip-gram Architecture1 Sep 2021 0 repositories listed
-
IRCologne at GermEval 2021: Toxicity Classification1 Sep 2021 0 repositories listed
-
Low-Resource ASR with an Augmented Language Model1 Sep 2021 0 repositories listed
-
Masked Adversarial Generation for Neural Machine Translation1 Sep 2021 0 repositories listed
-
Neural Borrowing Detection with Monolingual Lexical Models1 Sep 2021 0 repositories listed
-
On Reducing Repetition in Abstractive Summarization1 Sep 2021 0 repositories listed
-
Split-and-Rephrase in a Cross-Lingual Manner: A Complete Pipeline1 Sep 2021 0 repositories listed
-
Towards a Language Model for Temporal Commonsense Reasoning1 Sep 2021 0 repositories listed
-
Unsupervised Text Style Transfer with Content Embeddings1 Sep 2021 0 repositories listed
-
Watching a Language Model Learning Chess1 Sep 2021 0 repositories listed
-
Effectiveness of Deep Networks in NLP using BiDAF as an example architecture31 Aug 2021 0 repositories listed
-
How Does Adversarial Fine-Tuning Benefit BERT?31 Aug 2021 0 repositories listed
-
The effects of data size on Automated Essay Scoring engines30 Aug 2021 0 repositories listed
-
Representation Memorization for Fast Learning New Knowledge without Forgetting28 Aug 2021 0 repositories listed
-
Exploring Retraining-Free Speech Recognition for Intra-sentential Code-Switching27 Aug 2021 0 repositories listed
-
Exploring the Capacity of a Large-scale Masked Language Model to Recognize Grammatical Errors27 Aug 2021 0 repositories listed
-
Position-Invariant Truecasing with a Word-and-Character Hierarchical Recurrent Neural Network26 Aug 2021 0 repositories listed
-
Detection of Criminal Texts for the Polish State Border Guard24 Aug 2021 0 repositories listed
-
Prompt-Learning for Fine-Grained Entity Typing24 Aug 2021 0 repositories listed
-
Reducing Exposure Bias in Training Recurrent Neural Network Transducers24 Aug 2021 0 repositories listed
-
Using BERT Encoding and Sentence-Level Language Model for Sentence Ordering24 Aug 2021 0 repositories listed
-
UzBERT: pretraining a BERT model for Uzbek22 Aug 2021 0 repositories listed
-
Exploiting Multi-Object Relationships for Detecting Adversarial Attacks in Complex Scenes19 Aug 2021 0 repositories listed
-
Language Model Augmented Relevance Score19 Aug 2021 0 repositories listed
-
Automatic Multi-Label Prompting: Simple and Interpretable Few-Shot Classification17 Aug 2021 0 repositories listed
-
Deduplicating Training Data Makes Language Models Better17 Aug 2021 0 repositories listed
-
Scaling Laws for Deep Learning17 Aug 2021 0 repositories listed
-
Deep Natural Language Processing for LinkedIn Search16 Aug 2021 0 repositories listed
-
Towards Structured Dynamic Sparse Pre-Training of BERT13 Aug 2021 0 repositories listed
-
A Transformer-based Math Language Model for Handwritten Math Expression Recognition11 Aug 2021 0 repositories listed
-
Mounting Video Metadata on Transformer-based Language Model for Open-ended Video Question Answering11 Aug 2021 0 repositories listed
-
IntenT5: Search Result Diversification using Causal Language Models9 Aug 2021 0 repositories listed
-
Language Model Evaluation in Open-ended Text Generation8 Aug 2021 0 repositories listed
-
Leveraging Commonsense Knowledge on Classifying False News and Determining Checkworthiness of Claims8 Aug 2021 0 repositories listed
-
Sentence Semantic Regression for Text Generation6 Aug 2021 0 repositories listed
-
Towards Zero-shot Language Modeling6 Aug 2021 0 repositories listed
-
FMMformer: Efficient and Flexible Transformer via Decomposed Near-field and Far-field Attention5 Aug 2021 0 repositories listed
-
Mitigating harm in language models with conditional-likelihood filtration4 Aug 2021 0 repositories listed
-
Large-Scale Differentially Private BERT3 Aug 2021 0 repositories listed
-
Your fairness may vary: Pretrained language model fairness in toxic text classification3 Aug 2021 0 repositories listed
-
Is My Model Using The Right Evidence? Systematic Probes for Examining Evidence-Based Tabular Reasoning2 Aug 2021 0 repositories listed
-
A Comparison of Sentence-Weighting Techniques for NMT1 Aug 2021 0 repositories listed
-
A Targeted Assessment of Incremental Processing in Neural Language Models and Humans1 Aug 2021 0 repositories listed
-
AND does not mean OR: Using Formal Languages to Study Language Models' Representations1 Aug 2021 0 repositories listed
-
AStarTwice at SemEval-2021 Task 5: Toxic Span Detection Using RoBERTa-CRF, Domain Specific Pre-Training and Self-Training1 Aug 2021 0 repositories listed
-
Attending Self-Attention: A Case Study of Visually Grounded Supervision in Vision-and-Language Transformers1 Aug 2021 0 repositories listed