Browse State-of-the-Art › Language Modeling › Papers, page 113
Language Modeling
Papers archive 2025-07-28
archive papers tagged: 14,182 · with a code link: 5,620 · where Syntology ran a sample: 1,894 (1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,894 of 14,182 tagged: 1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument)
Page 113 of 142: papers 11,201 to 11,300 of 14,182, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Cross-Lingual QA as a Stepping Stone for Monolingual Open QA in Icelandic5 Jul 2022 0 repositories listed
-
MIA 2022 Shared Task Submission: Leveraging Entity Representations, Dense-Sparse Hybrids, and Fusion-in-Decoder for Cross-Lingual Question Answering5 Jul 2022 0 repositories listed
-
BERT, can HE predict contrastive focus? Predicting and controlling prominence in neural TTS using a language model4 Jul 2022 0 repositories listed
-
UserLibri: A Dataset for ASR Personalization Using Only Text2 Jul 2022 0 repositories listed
-
A Dog Is Passing Over The Jet? A Text-Generation Dataset for Korean Commonsense Reasoning and Evaluation1 Jul 2022 0 repositories listed
-
An Annotated Dataset and Automatic Approaches for Discourse Mode Identification in Low-resource Bengali Language1 Jul 2022 0 repositories listed
-
An Empirical Study on Pseudo-log-likelihood Bias Measures for Masked Language Models Using Paraphrased Sentences1 Jul 2022 0 repositories listed
-
AnaLog: Testing Analytical and Deductive Logic Learnability in Language Models1 Jul 2022 0 repositories listed
-
1 Jul 2022 0 repositories listed
-
DANGNT-SGU at SemEval-2022 Task 11: Using Pre-trained Language Model for Complex Named Entity Recognition1 Jul 2022 0 repositories listed
-
Data Augmentation with Dual Training for Offensive Span Detection1 Jul 2022 0 repositories listed
-
”Diversity and Uncertainty in Moderation” are the Key to Data Selection for Multilingual Few-shot Transfer1 Jul 2022 0 repositories listed
-
Don’t Forget About Pronouns: Removing Gender Bias in Language Models Without Losing Factual Gender Information1 Jul 2022 0 repositories listed
-
Empathetic Persuasion: Reinforcing Empathy and Persuasiveness in Dialogue Systems1 Jul 2022 0 repositories listed
-
Exploring the Effect of Dialect Mismatched Language Models in Telugu Automatic Speech Recognition1 Jul 2022 0 repositories listed
-
GPT-2-based Human-in-the-loop Theatre Play Script Generation1 Jul 2022 0 repositories listed
-
HuaAMS at SemEval-2022 Task 8: Combining Translation and Domain Pre-training for Cross-lingual News Article Similarity1 Jul 2022 0 repositories listed
-
Improving Classification of Infrequent Cognitive Distortions: Domain-Specific Model vs. Data Augmentation1 Jul 2022 0 repositories listed
-
Infrrd.ai at SemEval-2022 Task 11: A system for named entity recognition using data augmentation, transformer-based sequence labeling model, and EnsembleCRF1 Jul 2022 0 repositories listed
-
Intent Discovery for Enterprise Virtual Assistants: Applications of Utterance Embedding and Clustering to Intent Mining1 Jul 2022 0 repositories listed
-
JBNU-CCLab at SemEval-2022 Task 7: DeBERTa for Identifying Plausible Clarifications in Instructional Texts1 Jul 2022 0 repositories listed
-
KroneckerBERT: Significant Compression of Pre-trained Language Models Through Kronecker Decomposition and Knowledge Distillation1 Jul 2022 0 repositories listed
-
L3i at SemEval-2022 Task 11: Straightforward Additional Context for Multilingual Named Entity Recognition1 Jul 2022 0 repositories listed
-
Language Model Augmented Monotonic Attention for Simultaneous Translation1 Jul 2022 0 repositories listed
-
Mask and Regenerate: A Classifier-based Approach for Unpaired Sentiment Transformation of Reviews for Electronic Commerce Websites.1 Jul 2022 0 repositories listed
-
Masking Morphosyntactic Categories to Evaluate Salience for Schizophrenia Diagnosis1 Jul 2022 0 repositories listed
-
Minimally-Supervised Relation Induction from Pre-trained Language Model1 Jul 2022 0 repositories listed
-
MT-Speech at SemEval-2022 Task 10: Incorporating Data Augmentation and Auxiliary Task with Cross-Lingual Pretrained Language Model for Structured Sentiment Analysis1 Jul 2022 0 repositories listed
-
niksss at SemEval-2022 Task 6: Are Traditionally Pre-Trained Contextual Embeddings Enough for Detecting Intended Sarcasm ?1 Jul 2022 0 repositories listed
-
Self-supervised Product Title Rewrite for Product Listing Ads1 Jul 2022 0 repositories listed
-
SPDB Innovation Lab at SemEval-2022 Task 10: A Novel End-to-End Structured Sentiment Analysis Model based on the ERNIE-M1 Jul 2022 0 repositories listed
-
SwahBERT: Language Model of Swahili1 Jul 2022 0 repositories listed
-
TUG-CIC at SemEval-2021 Task 6: Two-stage Fine-tuning for Intended Sarcasm Detection1 Jul 2022 0 repositories listed
-
Uncertainty and Inclusivity in Gender Bias Annotation: An Annotation Taxonomy and Annotated Datasets of British English Text1 Jul 2022 0 repositories listed
-
Unsupervised Paraphrasability Prediction for Compound Nominalizations1 Jul 2022 0 repositories listed
-
Zuo Zhuan Ancient Chinese Dataset for Word Sense Disambiguation1 Jul 2022 0 repositories listed
-
"Diversity and Uncertainty in Moderation" are the Key to Data Selection for Multilingual Few-shot Transfer30 Jun 2022 0 repositories listed
-
Language model compression with weighted low-rank factorization30 Jun 2022 0 repositories listed
-
Contextual Density Ratio for Language Model Biasing of Sequence to Sequence ASR Systems29 Jun 2022 0 repositories listed
-
GPTs at Factify 2022: Prompt Aided Fact-Verification29 Jun 2022 0 repositories listed
-
Improving Deliberation by Text-Only and Semi-Supervised Training29 Jun 2022 0 repositories listed
-
Knowledge Distillation of Transformer-based Language Models Revisited29 Jun 2022 0 repositories listed
-
Simple and Effective Multi-sentence TTS with Expressive and Coherent Prosody29 Jun 2022 0 repositories listed
-
Adaptive Multi-view Rule Discovery for Weakly-Supervised Compatible Products Prediction28 Jun 2022 0 repositories listed
-
A Zero-Shot Classification Approach for a Word-Guessing Challenge27 Jun 2022 0 repositories listed
-
Bi-VLDoc: Bidirectional Vision-Language Modeling for Visually-Rich Document Understanding27 Jun 2022 0 repositories listed
-
Your Autoregressive Generative Model Can be Better If You Treat It as an Energy-Based One26 Jun 2022 0 repositories listed
-
Construct a Sentence with Multiple Specified Words25 Jun 2022 0 repositories listed
-
DP-Parse: Finding Word Boundaries from Raw Speech with an Instance Lexicon22 Jun 2022 0 repositories listed
-
Efficient and effective training of language and graph neural network models22 Jun 2022 0 repositories listed
-
Revisiting Group Differences in High-Dimensional Choices: Method and Application to Congressional Speech22 Jun 2022 0 repositories listed
-
Don't Forget About Pronouns: Removing Gender Bias in Language Models Without Losing Factual Gender Information21 Jun 2022 0 repositories listed
-
General Framework for Reversible Data Hiding in Texts Based on Masked Language Modeling21 Jun 2022 0 repositories listed
-
Knowledge Graph Fusion for Language Model Fine-tuning21 Jun 2022 0 repositories listed
-
TAPHSIR: Towards AnaPHoric Ambiguity Detection and ReSolution In Requirements21 Jun 2022 0 repositories listed
-
Using cognitive psychology to understand GPT-321 Jun 2022 0 repositories listed
-
Can Language Models Capture Graph Semantics? From Graphs to Language Model and Vice-Versa18 Jun 2022 0 repositories listed
-
Evolution through Large Models17 Jun 2022 0 repositories listed
-
Making first order linear logic a generating grammar17 Jun 2022 0 repositories listed
-
Accelerating Inference and Language Model Fusion of Recurrent Neural Network Transducers via End-to-End 4-bit Quantization16 Jun 2022 0 repositories listed
-
Characteristics of Harmful Text: Towards Rigorous Benchmarking of Language Models16 Jun 2022 0 repositories listed
-
Deep Multi-Task Models for Misogyny Identification and Categorization on Arabic Social Media16 Jun 2022 0 repositories listed
-
A Language Model With Million Context Length For Raw Audio16 Jun 2022 0 repositories listed
-
A Survey : Neural Networks for AMR-to-Text15 Jun 2022 0 repositories listed
-
Residual Language Model for End-to-end Speech Recognition15 Jun 2022 0 repositories listed
-
Test-Time Adaptation for Visual Document Understanding15 Jun 2022 0 repositories listed
-
Improving Pre-trained Language Model Fine-tuning with Noise Stability Regularization12 Jun 2022 0 repositories listed
-
Sort by Structure: Language Model Ranking as Dependency Probing10 Jun 2022 0 repositories listed
-
Context-based out-of-vocabulary word recovery for ASR systems in Indian languages9 Jun 2022 0 repositories listed
-
Learning to Generate Prompts for Dialogue Generation through Reinforcement Learning8 Jun 2022 0 repositories listed
-
DynaMaR: Dynamic Prompt with Mask Token Representation7 Jun 2022 0 repositories listed
-
ID-Agnostic User Behavior Pre-training for Sequential Recommendation6 Jun 2022 0 repositories listed
-
6 Jun 2022 0 repositories listed
-
A Control Theoretic Framework for Adaptive Gradient Optimizers in Machine Learning4 Jun 2022 0 repositories listed
-
Automatic Generation of Programming Exercises and Code Explanations using Large Language Models3 Jun 2022 0 repositories listed
-
Visual Clues: Bridging Vision and Language Foundations for Image Paragraph Captioning3 Jun 2022 0 repositories listed
-
BayesFormer: Transformer with Uncertainty Estimation2 Jun 2022 0 repositories listed
-
Code Generation Tools (Almost) for Free? A Study of Few-Shot, Pre-Trained Language Models on Code2 Jun 2022 0 repositories listed
-
VL-BEiT: Generative Vision-Language Pretraining2 Jun 2022 0 repositories listed
-
A Language Modelling Approach to Quality Assessment of OCR’ed Historical Text1 Jun 2022 0 repositories listed
-
Automatic Word Segmentation and Part-of-Speech Tagging of Ancient Chinese Based on BERT Model1 Jun 2022 0 repositories listed
-
BanglaHateBERT: BERT for Abusive Language Detection in Bengali1 Jun 2022 0 repositories listed
-
BEA-Base: A Benchmark for ASR of Spontaneous Hungarian1 Jun 2022 0 repositories listed
-
CHILLAX - at Arabic Hate Speech 2022: A Hybrid Machine Learning and Transformers based Model to Detect Arabic Offensive and Hate Speech1 Jun 2022 0 repositories listed
-
Clarifying Implicit and Underspecified Phrases in Instructional Text1 Jun 2022 0 repositories listed
-
Conversational Speech Recognition Needs Data? Experiments with Austrian German1 Jun 2022 0 repositories listed
-
CxLM: A Construction and Context-aware Language Model1 Jun 2022 0 repositories listed
-
Data Augmentation for Low-resource Word Segmentation and POS Tagging of Ancient Chinese Texts1 Jun 2022 0 repositories listed
-
Data Augmentation for the Post-Stroke Speech Transcription (PSST) Challenge: Sometimes Less Is More1 Jun 2022 0 repositories listed
-
Developing Language Resources and NLP Tools for the North Korean Language1 Jun 2022 0 repositories listed
-
Development and Evaluation of Speech Recognition for the Welsh Language1 Jun 2022 0 repositories listed
-
Discovering Financial Hypernyms by Prompting Masked Language Models1 Jun 2022 0 repositories listed
-
Do Transformer Networks Improve the Discovery of Rules from Text?1 Jun 2022 0 repositories listed
-
Efficiently and Thoroughly Anonymizing a Transformer Language Model for Dutch Electronic Health Records: a Two-Step Method1 Jun 2022 0 repositories listed
-
Electoral Agitation Dataset: The Use Case of the Polish Election1 Jun 2022 0 repositories listed
-
ENRICH4ALL: A First Luxembourgish BERT Model for a Multilingual Chatbot1 Jun 2022 0 repositories listed
-
Enriching Epidemiological Thematic Features For Disease Surveillance Corpora Classification1 Jun 2022 0 repositories listed
-
Error Correction Environment for the Polish Parliamentary Corpus1 Jun 2022 0 repositories listed
-
Evaluating Pre-Trained Language Models for Focused Terminology Extraction from Swedish Medical Records1 Jun 2022 0 repositories listed
-
Evaluating Pretraining Strategies for Clinical BERT Models1 Jun 2022 0 repositories listed