Browse State-of-the-Art › Language Modelling › Papers, page 147
Language Modelling
Papers archive 2025-07-28
archive papers tagged: 17,610 · with a code link: 7,012 · where Syntology ran a sample: 2,428 (2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,428 of 17,610 tagged: 2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument)
Page 147 of 177: papers 14,601 to 14,700 of 17,610, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
PINGAN Omini-Sinitic at SemEval-2021 Task 4:Reading Comprehension of Abstract Meaning1 Aug 2021 0 repositories listed
-
PRAL: A Tailored Pre-Training Model for Task-Oriented Dialog Generation1 Aug 2021 0 repositories listed
-
Probing Multi-modal Machine Translation with Pre-trained Language Model1 Aug 2021 0 repositories listed
-
Product Review Translation: Parallel Corpus Creation and Robustness towards User-generated Noisy Text1 Aug 2021 0 repositories listed
-
QASR: QCRI Aljazeera Speech Resource A Large Scale Annotated Arabic Speech Corpus1 Aug 2021 0 repositories listed
-
Rakuten’s Participation in WAT 2021: Examining the Effectiveness of Pre-trained Models for Multilingual and Multimodal Machine Translation1 Aug 2021 0 repositories listed
-
Realised Volatility Forecasting: Machine Learning via Financial Word Embedding1 Aug 2021 0 repositories listed
-
RoMa at SemEval-2021 Task 7: A Transformer-based Approach for Detecting and Rating Humor and Offense1 Aug 2021 0 repositories listed
-
S-NLP at SemEval-2021 Task 5: An Analysis of Dual Networks for Sequence Tagging1 Aug 2021 0 repositories listed
-
Selecting Informative Contexts Improves Language Model Fine-tuning1 Aug 2021 0 repositories listed
-
SkoltechNLP at SemEval-2021 Task 2: Generating Cross-Lingual Training Data for the Word-in-Context Task1 Aug 2021 0 repositories listed
-
Small-Scale Cross-Language Authorship Attribution on Social Media Comments1 Aug 2021 0 repositories listed
-
Stereotyping Norwegian Salmon: An Inventory of Pitfalls in Fairness Benchmark Datasets1 Aug 2021 0 repositories listed
-
Team “NoConflict” at CASE 2021 Task 1: Pretraining for Sentence-Level Protest Event Detection1 Aug 2021 0 repositories listed
-
The University of Edinburgh’s Submission to the IWSLT21 Simultaneous Translation Task1 Aug 2021 0 repositories listed
-
Unleash GPT-2 Power for Event Detection1 Aug 2021 0 repositories listed
-
Using Gender- and Polarity-Informed Models to Investigate Bias1 Aug 2021 0 repositories listed
-
Towards Continual Entity Learning in Language Models for Conversational Agents30 Jul 2021 0 repositories listed
-
Combining Probabilistic Logic and Deep Learning for Self-Supervised Learning27 Jul 2021 0 repositories listed
-
Cross-lingual Transferring of Pre-trained Contextualized Language Models27 Jul 2021 0 repositories listed
-
Exploiting Language Model for Efficient Linguistic Steganalysis26 Jul 2021 0 repositories listed
-
A Differentiable Language Model Adversarial Attack on Text Classifiers23 Jul 2021 0 repositories listed
-
Back-Translated Task Adaptive Pretraining: Improving Accuracy and Robustness on Text Classification22 Jul 2021 0 repositories listed
-
DeepTitle -- Leveraging BERT to generate Search Engine Optimized Headlines22 Jul 2021 0 repositories listed
-
Neuradicon: operational representation learning of neuroimaging reports21 Jul 2021 0 repositories listed
-
The Effectiveness of Intermediate-Task Training for Code-Switched Natural Language Understanding21 Jul 2021 0 repositories listed
-
Learning ULMFiT and Self-Distillation with Calibration for Medical Dialogue System20 Jul 2021 0 repositories listed
-
Seed Words Based Data Selection for Language Model Adaptation20 Jul 2021 0 repositories listed
-
Bridging the Gap between Language Model and Reading Comprehension: Unsupervised MRC via Self-Supervision19 Jul 2021 0 repositories listed
-
A Vector-Based Approach to Few-Shot Veracity Classification for Automated Fact-Checking17 Jul 2021 0 repositories listed
-
Using Language Models on Low-end Hardware17 Jul 2021 0 repositories listed
-
Are Multilingual Models the Best Choice for Moderately Under-resourced Languages? A Comprehensive Assessment for Catalan16 Jul 2021 0 repositories listed
-
Intersectional Bias in Causal Language Models16 Jul 2021 0 repositories listed
-
15 Jul 2021 0 repositories listed
-
DeepMutants: Training neural bug detectors with contextual mutations14 Jul 2021 0 repositories listed
-
From Show to Tell: A Survey on Deep Learning-based Image Captioning14 Jul 2021 0 repositories listed
-
14 Jul 2021 0 repositories listed
-
Large-Scale News Classification using BERT Language Model: Spark NLP Approach14 Jul 2021 0 repositories listed
-
Controlled Caption Generation for Images Through Adversarial Attacks7 Jul 2021 0 repositories listed
-
KOALA: A Kalman Optimization Algorithm with Loss Adaptivity7 Jul 2021 0 repositories listed
-
LanguageRefer: Spatial-Language Model for 3D Visual Grounding7 Jul 2021 0 repositories listed
-
Not Quite 'Ask a Librarian': AI on the Nature, Value, and Future of LIS7 Jul 2021 0 repositories listed
-
Cross-Lingual Transfer Learning for Statistical Type Inference1 Jul 2021 0 repositories listed
-
Getting to Production with Few-shot Natural Language Generation Models1 Jul 2021 0 repositories listed
-
Projection of Turn Completion in Incremental Spoken Dialogue Systems1 Jul 2021 0 repositories listed
-
Word-Free Spoken Language Understanding for Mandarin-Chinese1 Jul 2021 0 repositories listed
-
A Simple and Efficient Probabilistic Language model for Code-Mixed Text29 Jun 2021 0 repositories listed
-
A Knowledge-Grounded Dialog System Based on Pre-Trained Language Models28 Jun 2021 0 repositories listed
-
What's in a Measurement? Using GPT-3 on SemEval 2021 Task 8 -- MeasEval28 Jun 2021 0 repositories listed
-
Visual Conceptual Blending with Large-scale Language and Vision Models27 Jun 2021 0 repositories listed
-
Toward Less Hidden Cost of Code Completion with Acceptance and Ranking Models26 Jun 2021 0 repositories listed
-
Language Models are Good Translators25 Jun 2021 0 repositories listed
-
Learning to Sample Replacements for ELECTRA Pre-Training25 Jun 2021 0 repositories listed
-
25 Jun 2021 0 repositories listed
-
QASR: QCRI Aljazeera Speech Resource -- A Large Scale Annotated Arabic Speech Corpus24 Jun 2021 0 repositories listed
-
Winner Team Mia at TextVQA Challenge 2021: Vision-and-Language Representation Learning with Pre-trained Sequence-to-Sequence Model24 Jun 2021 0 repositories listed
-
CharacterChat: Supporting the Creation of Fictional Characters through Conversation and Progressive Manifestation with a Chatbot23 Jun 2021 0 repositories listed
-
Clinical Named Entity Recognition using Contextualized Token Representations23 Jun 2021 0 repositories listed
-
A Case Study in Bootstrapping Ontology Graphs from Textbooks22 Jun 2021 0 repositories listed
-
Data Augmentation for Opcode Sequence Based Malware Detection22 Jun 2021 0 repositories listed
-
Structured in Space, Randomized in Time: Leveraging Dropout in RNNs for Efficient Training22 Jun 2021 0 repositories listed
-
A Discriminative Entity-Aware Language Model for Virtual Assistants21 Jun 2021 0 repositories listed
-
Ad Text Classification with Transformer-Based Natural Language Processing Methods21 Jun 2021 0 repositories listed
-
Membership Inference on Word Embedding and Beyond21 Jun 2021 0 repositories listed
-
MaxUp: Lightweight Adversarial Training With Data Augmentation Improves Neural Network Training19 Jun 2021 0 repositories listed
-
19 Jun 2021 0 repositories listed
-
An Improved Single Step Non-autoregressive Transformer for Automatic Speech Recognition18 Jun 2021 0 repositories listed
-
Label prompt for multi-label text classification18 Jun 2021 0 repositories listed
-
Learning to Complete Code with Sketches18 Jun 2021 0 repositories listed
-
Low Resource German ASR with Untranscribed Data Spoken by Non-native Children -- INTERSPEECH 2021 Shared Task SPAPL System18 Jun 2021 0 repositories listed
-
Process for Adapting Language Models to Society (PALMS) with Values-Targeted Datasets18 Jun 2021 0 repositories listed
-
Algorithm to Compilation Co-design: An Integrated View of Neural Network Sparsity16 Jun 2021 0 repositories listed
-
Augmented Neural Story Generation with Commonsense Inference16 Jun 2021 0 repositories listed
-
SEOVER: Sentence-level Emotion Orientation Vector based Conversation Emotion Recognition Model16 Jun 2021 0 repositories listed
-
ASR Adaptation for E-commerce Chatbots using Cross-Utterance Context and Multi-Task Language Modeling15 Jun 2021 0 repositories listed
-
Dialectal Speech Recognition and Translation of Swiss German Speech to Standard German Text: Microsoft's Submission to SwissText 202115 Jun 2021 0 repositories listed
-
PairConnect: A Compute-Efficient MLP Alternative to Attention15 Jun 2021 0 repositories listed
-
Differentiable Neural Architecture Search with Morphism-based Transformable Backbone Architectures14 Jun 2021 0 repositories listed
-
Is Einstein more agreeable and less neurotic than Hitler? A computational exploration of the emotional and personality profiles of historical persons14 Jun 2021 0 repositories listed
-
Overcoming Domain Mismatch in Low Resource Sequence-to-Sequence ASR Models using Hybrid Generated Pseudotranscripts14 Jun 2021 0 repositories listed
-
Cross-utterance Reranking Models with BERT and Graph Convolutional Networks for Conversational Speech Recognition13 Jun 2021 0 repositories listed
-
Predicting the Ordering of Characters in Japanese Historical Documents12 Jun 2021 0 repositories listed
-
Leveraging Pre-trained Language Model for Speech Sentiment Analysis11 Jun 2021 0 repositories listed
-
Balanced End-to-End Monolingual pre-training for Low-Resourced Indic Languages Code-Switching Speech Recognition10 Jun 2021 0 repositories listed
-
MST: Masked Self-Supervised Transformer for Visual Representation10 Jun 2021 0 repositories listed
-
DGA-Net Dynamic Gaussian Attention Network for Sentence Semantic Matching9 Jun 2021 0 repositories listed
-
Hash Layers For Large Sparse Models8 Jun 2021 0 repositories listed
-
Measuring and Improving BERT's Mathematical Abilities by Predicting the Order of Reasoning7 Jun 2021 0 repositories listed
-
Pre-trained Language Model for Web-scale Retrieval in Baidu Search7 Jun 2021 0 repositories listed
-
RoSearch: Search for Robust Student Architectures When Distilling Pre-trained Language Models7 Jun 2021 0 repositories listed
-
Video Imprint7 Jun 2021 0 repositories listed
-
Let's be explicit about that: Distant supervision for implicit discourse relation classification via connective prediction6 Jun 2021 0 repositories listed
-
On the Effectiveness of Adapter-based Tuning for Pretrained Language Model Adaptation6 Jun 2021 0 repositories listed
-
Semantic-Enhanced Explainable Finetuning for Open-Domain Dialogues6 Jun 2021 0 repositories listed
-
Extracting Weighted Automata for Approximate Minimization in Language Modelling5 Jun 2021 0 repositories listed
-
Bi-Granularity Contrastive Learning for Post-Training in Few-Shot Scene4 Jun 2021 0 repositories listed
-
Exposing the Implicit Energy Networks behind Masked Language Models via Metropolis--Hastings4 Jun 2021 0 repositories listed
-
Language Model Metrics and Procrustes Analysis for Improved Vector Transformation of NLP Embeddings4 Jun 2021 0 repositories listed
-
Minimum Word Error Rate Training with Language Model Fusion for End-to-End Speech Recognition4 Jun 2021 0 repositories listed
-
nmT5 -- Is parallel data still relevant for pre-training massively multilingual language models?3 Jun 2021 0 repositories listed