Browse State-of-the-Art › Language Modelling › Papers, page 138
Language Modelling
Papers archive 2025-07-28
archive papers tagged: 17,610 · with a code link: 7,012 · where Syntology ran a sample: 2,428 (2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,428 of 17,610 tagged: 2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument)
Page 138 of 177: papers 13,701 to 13,800 of 17,610, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Automatic Word Segmentation and Part-of-Speech Tagging of Ancient Chinese Based on BERT Model1 Jun 2022 0 repositories listed
-
BanglaHateBERT: BERT for Abusive Language Detection in Bengali1 Jun 2022 0 repositories listed
-
BEA-Base: A Benchmark for ASR of Spontaneous Hungarian1 Jun 2022 0 repositories listed
-
CHILLAX - at Arabic Hate Speech 2022: A Hybrid Machine Learning and Transformers based Model to Detect Arabic Offensive and Hate Speech1 Jun 2022 0 repositories listed
-
Clarifying Implicit and Underspecified Phrases in Instructional Text1 Jun 2022 0 repositories listed
-
Combination of Contextualized and Non-Contextualized Layers for Lexical Substitution in French1 Jun 2022 0 repositories listed
-
Conversational Speech Recognition Needs Data? Experiments with Austrian German1 Jun 2022 0 repositories listed
-
CxLM: A Construction and Context-aware Language Model1 Jun 2022 0 repositories listed
-
Data Augmentation for Low-resource Word Segmentation and POS Tagging of Ancient Chinese Texts1 Jun 2022 0 repositories listed
-
Data Augmentation for the Post-Stroke Speech Transcription (PSST) Challenge: Sometimes Less Is More1 Jun 2022 0 repositories listed
-
Developing Language Resources and NLP Tools for the North Korean Language1 Jun 2022 0 repositories listed
-
Development and Evaluation of Speech Recognition for the Welsh Language1 Jun 2022 0 repositories listed
-
Discovering Financial Hypernyms by Prompting Masked Language Models1 Jun 2022 0 repositories listed
-
Do Transformer Networks Improve the Discovery of Rules from Text?1 Jun 2022 0 repositories listed
-
Efficiently and Thoroughly Anonymizing a Transformer Language Model for Dutch Electronic Health Records: a Two-Step Method1 Jun 2022 0 repositories listed
-
Electoral Agitation Dataset: The Use Case of the Polish Election1 Jun 2022 0 repositories listed
-
ENRICH4ALL: A First Luxembourgish BERT Model for a Multilingual Chatbot1 Jun 2022 0 repositories listed
-
Enriching Epidemiological Thematic Features For Disease Surveillance Corpora Classification1 Jun 2022 0 repositories listed
-
Error Correction Environment for the Polish Parliamentary Corpus1 Jun 2022 0 repositories listed
-
Evaluating Pre-Trained Language Models for Focused Terminology Extraction from Swedish Medical Records1 Jun 2022 0 repositories listed
-
Evaluating Pretraining Strategies for Clinical BERT Models1 Jun 2022 0 repositories listed
-
Evaluating Unsupervised Approaches to Morphological Segmentation for Wolastoqey1 Jun 2022 0 repositories listed
-
FQuAD2.0: French Question Answering and Learning When You Don’t Know1 Jun 2022 0 repositories listed
-
From FreEM to D’AlemBERT: a Large Corpus and a Language Model for Early Modern French1 Jun 2022 0 repositories listed
-
gaBERT — an Irish Language Model1 Jun 2022 0 repositories listed
-
HADREB: Human Appraisals and (English) Descriptions of Robot Emotional Behaviors1 Jun 2022 0 repositories listed
-
HerBERT Based Language Model Detects Quantifiers and Their Semantic Properties in Polish1 Jun 2022 0 repositories listed
-
Impact Analysis of the Use of Speech and Language Models Pretrained by Self-Supersivion for Spoken Language Understanding1 Jun 2022 0 repositories listed
-
Korean Language Modeling via Syntactic Guide1 Jun 2022 0 repositories listed
-
LARSA22 at Qur’an QA 2022: Text-to-Text Transformer for Finding Answers to Questions from Qur’an1 Jun 2022 0 repositories listed
-
Latvian National Corpora Collection – Korpuss.lv1 Jun 2022 0 repositories listed
-
Lessons Learned from GPT-SW3: Building the First Large-Scale Generative Language Model for Swedish1 Jun 2022 0 repositories listed
-
LuxemBERT: Simple and Practical Data Augmentation in Language Model Pre-Training for Luxembourgish1 Jun 2022 0 repositories listed
-
1 Jun 2022 0 repositories listed
-
Modeling Dutch Medical Texts for Detecting Functional Categories and Levels of COVID-19 Patients1 Jun 2022 0 repositories listed
-
Multilingual and Multimodal Learning for Brazilian Portuguese1 Jun 2022 0 repositories listed
-
Multilingual Comparative Analysis of Deep-Learning Dependency Parsing Results Using Parallel Corpora1 Jun 2022 0 repositories listed
-
Nepali Encoder Transformers: An Analysis of Auto Encoding Transformer Language Models for Nepali Text Classification1 Jun 2022 0 repositories listed
-
Sentiment Analysis of Homeric Text: The 1st Book of Iliad1 Jun 2022 0 repositories listed
-
Sign Language Production With Avatar Layering: A Critical Use Case over Rare Words1 Jun 2022 0 repositories listed
-
Simple Tagging System with RoBERTa for Ancient Chinese1 Jun 2022 0 repositories listed
-
SPOCK at FinCausal 2022: Causal Information Extraction Using Span-Based and Sequence Tagging Models1 Jun 2022 0 repositories listed
-
Towards an Open-Source Dutch Speech Recognition System for the Healthcare Domain1 Jun 2022 0 repositories listed
-
Towards the Detection of a Semantic Gap in the Chain of Commonsense Knowledge Triples1 Jun 2022 0 repositories listed
-
Tracking Changes in ESG Representation: Initial Investigations in UK Annual Reports1 Jun 2022 0 repositories listed
-
Transformer with Fourier Integral Attentions1 Jun 2022 0 repositories listed
-
Word Class Based Language Modeling: A Case of Upper Sorbian1 Jun 2022 0 repositories listed
-
A Mixture-of-Expert Approach to RL-based Dialogue Management31 May 2022 0 repositories listed
-
The Contribution of Lyrics and Acoustics to Collaborative Understanding of Mood31 May 2022 0 repositories listed
-
COFS: Controllable Furniture layout Synthesis29 May 2022 0 repositories listed
-
Urdu News Article Recommendation Model using Natural Language Processing Techniques29 May 2022 0 repositories listed
-
Few-shot Subgoal Planning with Language Models28 May 2022 0 repositories listed
-
Happenstance: Utilizing Semantic Search to Track Russian State Media Narratives about the Russo-Ukrainian War On Reddit28 May 2022 0 repositories listed
-
Differentially Private Decoding in Large Language Models26 May 2022 0 repositories listed
-
Ground-Truth Labels Matter: A Deeper Look into Input-Label Demonstrations25 May 2022 0 repositories listed
-
Improving CTC-based ASR Models with Gated Interlayer Collaboration25 May 2022 0 repositories listed
-
Investigating Lexical Replacements for Arabic-English Code-Switched Data Augmentation25 May 2022 0 repositories listed
-
Know Where You're Going: Meta-Learning for Parameter-Efficient Fine-Tuning25 May 2022 0 repositories listed
-
Large Language Models are Few-Shot Clinical Information Extractors25 May 2022 0 repositories listed
-
Low Resource Style Transfer via Domain Adaptive Meta Learning25 May 2022 0 repositories listed
-
Segmenting Numerical Substitution Ciphers25 May 2022 0 repositories listed
-
Evaluating the Impact of Model Scale for Compositional Generalization in Semantic Parsing24 May 2022 0 repositories listed
-
Enhancing Continual Learning with Global Prototypes: Counteracting Negative Representation Drift24 May 2022 0 repositories listed
-
MaskEval: Weighted MLM-Based Evaluation for Text Summarization and Simplification24 May 2022 0 repositories listed
-
Multi-Level Modeling Units for End-to-End Mandarin Speech Recognition24 May 2022 0 repositories listed
-
On the Role of Bidirectionality in Language Model Pre-Training24 May 2022 0 repositories listed
-
Toxicity Detection with Generative Prompt-based Inference24 May 2022 0 repositories listed
-
Improving Short Text Classification With Augmented Data Using GPT-323 May 2022 0 repositories listed
-
RL with KL penalties is better viewed as Bayesian inference23 May 2022 0 repositories listed
-
The Diminishing Returns of Masked Language Models to Science23 May 2022 0 repositories listed
-
Supporting Vision-Language Model Inference with Causality-pruning Knowledge Prompt23 May 2022 0 repositories listed
-
Memorization Without Overfitting: Analyzing the Training Dynamics of Large Language Models22 May 2022 0 repositories listed
-
Named Entity Linking with Entity Representation by Multiple Embeddings21 May 2022 0 repositories listed
-
Scenario-based Multi-product Advertising Copywriting Generation for E-Commerce21 May 2022 0 repositories listed
-
Are Prompt-based Models Clueless?19 May 2022 0 repositories listed
-
Automatic Spoken Language Identification using a Time-Delay Neural Network19 May 2022 0 repositories listed
-
Foundation Posteriors for Approximate Probabilistic Inference19 May 2022 0 repositories listed
-
Great Power, Great Responsibility: Recommendations for Reducing Energy for Training Language Models19 May 2022 0 repositories listed
-
Minimising Biasing Word Errors for Contextual ASR with the Tree-Constrained Pointer Generator18 May 2022 0 repositories listed
-
Feature Aggregation in Zero-Shot Cross-Lingual Transfer Using Multilingual BERT17 May 2022 0 repositories listed
-
M6-Rec: Generative Pretrained Language Models are Open-Ended Recommender Systems17 May 2022 0 repositories listed
-
TiBERT: Tibetan Pre-trained Language Model15 May 2022 0 repositories listed
-
Controlling Translation Formality Using Pre-trained Multilingual Language Models13 May 2022 0 repositories listed
-
PathologyBERT -- Pre-trained Vs. A New Transformer Language Model for Pathology Domain13 May 2022 0 repositories listed
-
Efficient and Training-Free Control of Language Generation12 May 2022 0 repositories listed
-
Aggregating Pairwise Semantic Differences for Few-Shot Claim Veracity Classification11 May 2022 0 repositories listed
-
Sentence-level Privacy for Document Embeddings10 May 2022 0 repositories listed
-
The Importance of Context in Very Low Resource Language Modeling10 May 2022 0 repositories listed
-
LayoutXLM vs. GNN: An Empirical Evaluation of Relation Extraction for Documents9 May 2022 0 repositories listed
-
Multi-segment preserving sampling for deep manifold sampler9 May 2022 0 repositories listed
-
Towards a Progression-Aware Autonomous Dialogue Agent7 May 2022 0 repositories listed
-
Prompt Distribution Learning6 May 2022 0 repositories listed
-
Assistive Recipe Editing through Critiquing5 May 2022 0 repositories listed
-
Explain and Conquer: Personalised Text-based Reviews to Achieve Transparency3 May 2022 0 repositories listed
-
SparCAssist: A Model Risk Assessment Assistant Based on Sparse Generated Counterfactuals3 May 2022 0 repositories listed
-
A Holistic Assessment of the Carbon Footprint of Noor, a Very Large Arabic Language Model1 May 2022 0 repositories listed
-
A Knowledge storage and semantic space alignment Method for Multi-documents dialogue generation1 May 2022 0 repositories listed
-
Adaptive Differential Privacy for Language Model Training1 May 2022 0 repositories listed
-
Adversarial Soft Prompt Tuning for Cross-Domain Sentiment Analysis1 May 2022 0 repositories listed
-
AlephBERT: Language Model Pre-training and Evaluation from Sub-Word to Sentence Level1 May 2022 0 repositories listed