Browse State-of-the-Art › Language Modelling › Papers, page 148
Language Modelling
Papers archive 2025-07-28
archive papers tagged: 17,610 · with a code link: 7,012 · where Syntology ran a sample: 2,428 (2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,428 of 17,610 tagged: 2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument)
Page 148 of 177: papers 14,701 to 14,800 of 17,610, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
The Limitations of Limited Context for Constituency Parsing3 Jun 2021 0 repositories listed
-
A Span Extraction Approach for Information Extraction on Visually-Rich Documents2 Jun 2021 0 repositories listed
-
belabBERT: a Dutch RoBERTa-based language model applied to psychiatric classification2 Jun 2021 0 repositories listed
-
Learning to Select: A Fully Attentive Approach for Novel Object Captioning2 Jun 2021 0 repositories listed
-
One Teacher is Enough? Pre-trained Language Model Distillation from Multiple Teachers2 Jun 2021 0 repositories listed
-
Ad Headline Generation using Self-Critical Masked Language Model1 Jun 2021 0 repositories listed
-
Adversities are all you need: Classification of self-reported breast cancer posts on Twitter using Adversarial Fine-tuning1 Jun 2021 0 repositories listed
-
Are we there yet? Exploring clinical domain knowledge of BERT models1 Jun 2021 0 repositories listed
-
Classification, Extraction, and Normalization : CASIA_Unisound Team at the Social Media Mining for Health 2021 Shared Tasks1 Jun 2021 0 repositories listed
-
Classification of Tweets Self-reporting Adverse Pregnancy Outcomes and Potential COVID-19 Cases Using RoBERTa Transformers1 Jun 2021 0 repositories listed
-
DamascusTeam at NLP4IF2021: Fighting the Arabic COVID-19 Infodemic on Twitter Using AraBERT1 Jun 2021 0 repositories listed
-
ERNIE-NLI: Analyzing the Impact of Domain-Specific External Knowledge on Enhanced Representations for NLI1 Jun 2021 0 repositories listed
-
Explainable Multi-hop Verbal Reasoning Through Internal Monologue1 Jun 2021 0 repositories listed
-
FlowPrior: Learning Expressive Priors for Latent Variable Sentence Models1 Jun 2021 0 repositories listed
-
Identification de profil clinique du patient: Une approche de classification de séquences utilisant des modèles de langage français contextualisés (Identification of patient clinical profiles : A sequence classification approach using contextualised French language models )1 Jun 2021 0 repositories listed
-
Learning and Evaluating a Differentially Private Pre-trained Language Model1 Jun 2021 0 repositories listed
-
LongSumm 2021: Session based automatic summarization model for scientific document1 Jun 2021 0 repositories listed
-
Low-Resource Machine Translation Using Cross-Lingual Language Model Pretraining1 Jun 2021 0 repositories listed
-
MG-BERT: Multi-Graph Augmented BERT for Masked Language Modeling1 Jun 2021 0 repositories listed
-
Nutri-bullets Hybrid: Consensual Multi-document Summarization1 Jun 2021 0 repositories listed
-
On Randomized Classification Layers and Their Implications in Natural Language Generation1 Jun 2021 0 repositories listed
-
PIGLeT: Language Grounding Through Neuro-Symbolic Interaction in a 3D World1 Jun 2021 0 repositories listed
-
Predicting Numerals in Natural Language Text Using a Language Model Considering the Quantitative Aspects of Numerals1 Jun 2021 0 repositories listed
-
SCRIPT: Self-Critic PreTraining of Transformers1 Jun 2021 0 repositories listed
-
Structural Realization with GGNNs1 Jun 2021 0 repositories listed
-
System description for ProfNER - SMMH: Optimized finetuning of a pretrained transformer and word vectors1 Jun 2021 0 repositories listed
-
Target-Aware Data Augmentation for Stance Detection1 Jun 2021 0 repositories listed
-
Transferring Representations of Logical Connectives1 Jun 2021 0 repositories listed
-
Unsupervised Domain Adaptation in Cross-corpora Abusive Language Detection1 Jun 2021 0 repositories listed
-
Language Model Evaluation Beyond Perplexity31 May 2021 0 repositories listed
-
Picking Pearl From Seabed: Extracting Artefacts from Noisy Issue Triaging Collaborative Conversations for Hybrid Cloud Services31 May 2021 0 repositories listed
-
Text Summarization with Latent Queries31 May 2021 0 repositories listed
-
Tesseract: Parallelize the Tensor Parallelism Efficiently30 May 2021 0 repositories listed
-
NAS-BERT: Task-Agnostic and Adaptive-Size BERT Compression with Neural Architecture Search30 May 2021 0 repositories listed
-
Predictive Representation Learning for Language Modeling29 May 2021 0 repositories listed
-
Investigating Code-Mixed Modern Standard Arabic-Egyptian to English Machine Translation28 May 2021 0 repositories listed
-
Generative Adversarial Imitation Learning for Empathy-based AI27 May 2021 0 repositories listed
-
Leveraging Linguistic Coordination in Reranking N-Best Candidates For End-to-End Response Selection Using BERT27 May 2021 0 repositories listed
-
On Privacy and Confidentiality of Communications in Organizational Graphs27 May 2021 0 repositories listed
-
SGPT: Semantic Graphs based Pre-training for Aspect-based Sentiment Analysis26 May 2021 0 repositories listed
-
Towards an IMU-based Pen Online Handwriting Recognizer26 May 2021 0 repositories listed
-
NukeLM: Pre-Trained and Fine-Tuned Language Models for the Nuclear and Energy Domains25 May 2021 0 repositories listed
-
Pre-trained Language Model based Ranking in Baidu Search24 May 2021 0 repositories listed
-
LMSOC: An approach for socially sensitive pretraining22 May 2021 0 repositories listed
-
Aligning Visual Prototypes with BERT Embeddings for Few-Shot Learning21 May 2021 0 repositories listed
-
Unsupervised Multilingual Sentence Embeddings for Parallel Corpus Mining21 May 2021 0 repositories listed
-
See, Hear, Read: Leveraging Multimodality with Guided Attention for Abstractive Text Summarization20 May 2021 0 repositories listed
-
Accelerating Gossip SGD with Periodic Global Averaging19 May 2021 0 repositories listed
-
Exploring Text-to-Text Transformers for English to Hinglish Machine Translation with Synthetic Code-Mixing18 May 2021 0 repositories listed
-
Sentence Similarity Based on Contexts17 May 2021 0 repositories listed
-
Text based personality prediction from multiple social media data sources using pre-trained language model and model averaging17 May 2021 0 repositories listed
-
Neural Predictive Text for Grammatical Error Prevention16 May 2021 0 repositories listed
-
SINA-BERT: A Pre-Trained Language Model for Analysis of Medical Texts in Persian16 May 2021 0 repositories listed
-
A Cognitive Regularizer for Language Modeling15 May 2021 0 repositories listed
-
Towards Human-Free Automatic Quality Evaluation of German Summarization13 May 2021 0 repositories listed
-
Slower is Better: Revisiting the Forgetting Mechanism in LSTM for Slower Information Decay12 May 2021 0 repositories listed
-
10 May 2021 0 repositories listed
-
Computer-Aided Design as Language6 May 2021 0 repositories listed
-
HerBERT: Efficiently Pretrained Transformer-based Language Model for Polish4 May 2021 0 repositories listed
-
Goldilocks: Just-Right Tuning of BERT for Technology-Assisted Review3 May 2021 0 repositories listed
-
Impact of Gender Debiased Word Embeddings in Language Modeling3 May 2021 0 repositories listed
-
3 May 2021 0 repositories listed
-
Larger-Scale Transformers for Multilingual Masked Language Modeling2 May 2021 0 repositories listed
-
It’s Basically the Same Language Anyway: the Case for a Nordic Language Model1 May 2021 0 repositories listed
-
Measuring Translationese across Levels of Expertise: Are Professionals more Surprising than Students?1 May 2021 0 repositories listed
-
An analysis of full-size Russian complexly NER labelled corpus of Internet user reviews on the drugs based on deep learning and language neural nets30 Apr 2021 0 repositories listed
-
The Zero Resource Speech Challenge 2021: Spoken language modelling29 Apr 2021 0 repositories listed
-
FastAdaBelief: Improving Convergence Rate for Belief-based Adaptive Optimizers by Exploiting Strong Convexity28 Apr 2021 0 repositories listed
-
Extractive and Abstractive Explanations for Fact-Checking and Evaluation of News27 Apr 2021 0 repositories listed
-
Diverse Image Inpainting with Bidirectional and Autoregressive Transformers26 Apr 2021 0 repositories listed
-
Teaching a Massive Open Online Course on Natural Language Processing26 Apr 2021 0 repositories listed
-
Reranking Machine Translation Hypotheses with Structured and Web-based Language Models25 Apr 2021 0 repositories listed
-
BERT-CoQAC: BERT-based Conversational Question Answering in Context23 Apr 2021 0 repositories listed
-
Transfer training from smaller language model23 Apr 2021 0 repositories listed
-
Fast Text-Only Domain Adaptation of RNN-Transducer Prediction Network22 Apr 2021 0 repositories listed
-
Extracting Adverse Drug Events from Clinical Notes21 Apr 2021 0 repositories listed
-
On Sampling-Based Training Criteria for Neural Language Modeling21 Apr 2021 0 repositories listed
-
Pre-training for Spoken Language Understanding with Joint Textual and Phonetic Representation Learning21 Apr 2021 0 repositories listed
-
Adapting Long Context NLM for ASR Rescoring in Conversational Agents21 Apr 2021 0 repositories listed
-
19 Apr 2021 0 repositories listed
-
Understanding Chinese Video and Language via Contrastive Multimodal Pre-Training19 Apr 2021 0 repositories listed
-
On the Influence of Masking Policies in Intermediate Pre-training18 Apr 2021 0 repositories listed
-
16 Apr 2021 0 repositories listed
-
Enriching a Model's Notion of Belief using a Persistent Memory16 Apr 2021 0 repositories listed
-
15 Apr 2021 0 repositories listed
-
Natural Language Understanding with Privacy-Preserving BERT15 Apr 2021 0 repositories listed
-
15 Apr 2021 0 repositories listed
-
SINA-BERT: A pre-trained Language Model for Analysis of Medical Texts in Persian15 Apr 2021 0 repositories listed
-
Ultra-High Dimensional Sparse Representations with Binarization for Efficient Text Retrieval15 Apr 2021 0 repositories listed
-
Large-Scale Self- and Semi-Supervised Learning for Speech Translation14 Apr 2021 0 repositories listed
-
Masked Language Modeling and the Distributional Hypothesis: Order Word Matters Pre-training for Little14 Apr 2021 0 repositories listed
-
Mean-Squared Accuracy of Good-Turing Estimator14 Apr 2021 0 repositories listed
-
Experiments of ASR-based mispronunciation detection for children and adult English learners13 Apr 2021 0 repositories listed
-
Multilingual Transfer Learning for Code-Switched Language and Speech Neural Modeling13 Apr 2021 0 repositories listed
-
Restoring and Mining the Records of the Joseon Dynasty via Neural Language Modeling and Machine Translation13 Apr 2021 0 repositories listed
-
Should Semantic Vector Composition be Explicit? Can it be Linear?13 Apr 2021 0 repositories listed
-
Transformer-based Methods for Recognizing Ultra Fine-grained Entities (RUFES)13 Apr 2021 0 repositories listed
-
What's in your Head? Emergent Behaviour in Multi-Task Transformer Models13 Apr 2021 0 repositories listed
-
Comparing the Benefit of Synthetic Training Data for Various Automatic Speech Recognition Architectures12 Apr 2021 0 repositories listed
-
Estimating Subjective Crowd-Evaluations as an Additional Objective to Improve Natural Language Generation12 Apr 2021 0 repositories listed