Browse State-of-the-Art › Language Modelling › Papers, page 134
Language Modelling
Papers archive 2025-07-28
archive papers tagged: 17,610 · with a code link: 7,012 · where Syntology ran a sample: 2,428 (2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,428 of 17,610 tagged: 2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument)
Page 134 of 177: papers 13,301 to 13,400 of 17,610, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Leveraging per Image-Token Consistency for Vision-Language Pre-training20 Nov 2022 0 repositories listed
-
Leveraging Users' Social Network Embeddings for Fake News Detection on Twitter19 Nov 2022 0 repositories listed
-
3d human motion generation from the text via gesture action classification and the autoregressive model18 Nov 2022 0 repositories listed
-
Metadata Might Make Language Models Better18 Nov 2022 0 repositories listed
-
LongFNT: Long-form Speech Recognition with Factorized Neural Transducer17 Nov 2022 0 repositories listed
-
CAPE: Corrective Actions from Precondition Errors using Large Language Models17 Nov 2022 0 repositories listed
-
Prompting PaLM for Translation: Assessing Strategies and Performance16 Nov 2022 0 repositories listed
-
Technical Report on Neural Language Models and Few-Shot Learning for Systematic Requirements Processing in MDSE16 Nov 2022 0 repositories listed
-
Towards Computationally Verifiable Semantic Grounding for Language Models16 Nov 2022 0 repositories listed
-
TSMind: Alibaba and Soochow University's Submission to the WMT22 Translation Suggestion Task16 Nov 2022 0 repositories listed
-
ED-FAITH: Evaluating Dialogue Summarization on Faithfulness15 Nov 2022 0 repositories listed
-
Empowering Language Models with Knowledge Graph Reasoning for Question Answering15 Nov 2022 0 repositories listed
-
FedTune: A Deep Dive into Efficient Federated Fine-Tuning with Pre-trained Transformers15 Nov 2022 0 repositories listed
-
Introducing Semantics into Speech Encoders15 Nov 2022 0 repositories listed
-
Reasoning Circuits: Few-shot Multihop Question Generation with Structured Rationales15 Nov 2022 0 repositories listed
-
Relationship of the language distance to English ability of a country15 Nov 2022 0 repositories listed
-
RobBERT-2022: Updating a Dutch Language Model to Account for Evolving Language Use15 Nov 2022 0 repositories listed
-
ALBERT with Knowledge Graph Encoder Utilizing Semantic Similarity for Commonsense Question Answering14 Nov 2022 0 repositories listed
-
Grafting Pre-trained Models for Multimodal Headline Generation14 Nov 2022 0 repositories listed
-
Towards a Mathematics Formalisation Assistant using Large Language Models14 Nov 2022 0 repositories listed
-
Textual Data Augmentation for Patient Outcomes Prediction13 Nov 2022 0 repositories listed
-
DocuT5: Seq2seq SQL Generation with Table Documentation11 Nov 2022 0 repositories listed
-
BERT in Plutarch's Shadows10 Nov 2022 0 repositories listed
-
FormLM: Recommending Creation Ideas for Online Forms by Modelling Semantic and Structural Information10 Nov 2022 0 repositories listed
-
Prompt Learning for Domain Adaptation in Task-Oriented Dialogue10 Nov 2022 0 repositories listed
-
Syntax-Guided Domain Adaptation for Aspect-based Sentiment Analysis10 Nov 2022 0 repositories listed
-
The CRINGE Loss: Learning what language not to model10 Nov 2022 0 repositories listed
-
Adaptive Multi-Corpora Language Model Training for Speech Recognition9 Nov 2022 0 repositories listed
-
9 Nov 2022 0 repositories listed
-
FF2: A Feature Fusion Two-Stream Framework for Punctuation Restoration9 Nov 2022 0 repositories listed
-
Improving Noisy Student Training on Non-target Domain Data for Automatic Speech Recognition9 Nov 2022 0 repositories listed
-
Training self-supervised peptide sequence models on artificially chopped proteins9 Nov 2022 0 repositories listed
-
Understanding Cross-modal Interactions in V&L Models that Generate Scene Descriptions9 Nov 2022 0 repositories listed
-
Active Learning with Tabular Language Models8 Nov 2022 0 repositories listed
-
Parameter and Data Efficient Continual Pre-training for Robustness to Dialectal Variance in Arabic8 Nov 2022 0 repositories listed
-
Self-conditioned Embedding Diffusion for Text Generation8 Nov 2022 0 repositories listed
-
Word Order Matters when you Increase Masking8 Nov 2022 0 repositories listed
-
Complex Reading Comprehension Through Question Decomposition7 Nov 2022 0 repositories listed
-
7 Nov 2022 0 repositories listed
-
Prompter: Utilizing Large Language Model Prompting for a Data Efficient Embodied Instruction Following7 Nov 2022 0 repositories listed
-
Noisy Channel for Automatic Text Simplification6 Nov 2022 0 repositories listed
-
Measuring Progress on Scalable Oversight for Large Language Models4 Nov 2022 0 repositories listed
-
OSIC: A New One-Stage Image Captioner Coined4 Nov 2022 0 repositories listed
-
Circling Back to Recurrent Models of Language3 Nov 2022 0 repositories listed
-
Probing Statistical Representations For End-To-End ASR3 Nov 2022 0 repositories listed
-
Using Large Pre-Trained Language Model to Assist FDA in Premarket Medical Device3 Nov 2022 0 repositories listed
-
BECTRA: Transducer-based End-to-End ASR with BERT-Enhanced Encoder2 Nov 2022 0 repositories listed
-
Generative Adversarial Training Can Improve Neural Language Models2 Nov 2022 0 repositories listed
-
Internal Language Model Estimation based Adaptive Language Model Fusion for Domain Adaptation2 Nov 2022 0 repositories listed
-
Multi-level Distillation of Semantic Knowledge for Pre-training Multilingual Language Model2 Nov 2022 0 repositories listed
-
Numerical Optimizations for Weighted Low-rank Estimation on Language Model2 Nov 2022 0 repositories listed
-
Towards Zero-Shot Code-Switched Speech Recognition2 Nov 2022 0 repositories listed
-
A Quantitative Analysis of Comparison of Emoji Sentiment: Taiwan Mandarin Users and English Users1 Nov 2022 0 repositories listed
-
HanTrans: An Empirical Study on Cross-Era Transferability of Chinese Pre-trained Language Model1 Nov 2022 0 repositories listed
-
Language Model Based Chinese Handwriting Address Recognition1 Nov 2022 0 repositories listed
-
Machine learning can guide experimental approaches for protein digestibility estimations1 Nov 2022 0 repositories listed
-
NERVE at ROCLING 2022 Shared Task: A Comparison of Three Named Entity Recognition Frameworks Based on Language Model and Lexicon Approach1 Nov 2022 0 repositories listed
-
Two-stage LLM Fine-tuning with Less Specialization and More Generalization1 Nov 2022 0 repositories listed
-
Reduce, Reuse, Recycle: Improving Training Efficiency with Distillation1 Nov 2022 0 repositories listed
-
The future is different: Large pre-trained language models fail in prediction tasks1 Nov 2022 0 repositories listed
-
1 Nov 2022 0 repositories listed
-
A Simple, Yet Effective Approach to Finding Biases in Code Generation31 Oct 2022 0 repositories listed
-
Generating Sequences by Learning to Self-Correct31 Oct 2022 0 repositories listed
-
Modular Hybrid Autoregressive Transducer31 Oct 2022 0 repositories listed
-
Pneg: Prompt-based Negative Response Generation for Dialogue Response Selection Task31 Oct 2022 0 repositories listed
-
Tables to LaTeX: structure and content extraction from scientific tables31 Oct 2022 0 repositories listed
-
Learning to Decompose: Hypothetical Question Decomposition Based on Comparable Texts30 Oct 2022 0 repositories listed
-
token2vec: A Joint Self-Supervised Pre-training Framework Using Unpaired Speech and Text30 Oct 2022 0 repositories listed
-
BERT Meets CTC: New Formulation of End-to-End Speech Recognition with Pre-trained Masked Language Model29 Oct 2022 0 repositories listed
-
NTULM: Enriching Social Media Text Representations with Non-Textual Units29 Oct 2022 0 repositories listed
-
DiMBERT: Learning Vision-Language Grounded Representations with Disentangled Multimodal-Attention28 Oct 2022 0 repositories listed
-
Feature Engineering vs BERT on Twitter Data28 Oct 2022 0 repositories listed
-
28 Oct 2022 0 repositories listed
-
Modeling structure-building in the brain with CCG parsing and large language models28 Oct 2022 0 repositories listed
-
UPainting: Unified Text-to-Image Diffusion Generation with Cross-modal Guidance28 Oct 2022 0 repositories listed
-
Learning Joint Representation of Human Motion and Language27 Oct 2022 0 repositories listed
-
Nearest Neighbor Language Models for Stylistic Controllable Generation27 Oct 2022 0 repositories listed
-
SAN: a robust end-to-end ASR model architecture27 Oct 2022 0 repositories listed
-
Self-supervised language learning from raw audio: Lessons from the Zero Resource Speech Challenge27 Oct 2022 0 repositories listed
-
Seq2Seq-SC: End-to-End Semantic Communication Systems with Pre-trained Language Model27 Oct 2022 0 repositories listed
-
Simulating realistic speech overlaps improves multi-talker ASR27 Oct 2022 0 repositories listed
-
Bloom Library: Multimodal Datasets in 300+ Languages for a Variety of Downstream Tasks26 Oct 2022 0 repositories listed
-
Incorporating Pre-training Paradigm for Antibody Sequence-Structure Co-design26 Oct 2022 0 repositories listed
-
Piloting Copilot, Codex, and StarCoder2: Hot Temperature, Cold Prompts, or Black Magic?26 Oct 2022 0 repositories listed
-
Cloning Ideology and Style using Deep Learning25 Oct 2022 0 repositories listed
-
Differentially Private Language Models for Secure Data Sharing25 Oct 2022 0 repositories listed
-
Dual Mechanism Priming Effects in Hindi Word Order25 Oct 2022 0 repositories listed
-
Learning Better Intent Representations for Financial Open Intent Classification25 Oct 2022 0 repositories listed
-
Leveraging Open Data and Task Augmentation to Automated Behavioral Coding of Psychotherapy Conversations in Low-Resource Scenarios25 Oct 2022 0 repositories listed
-
Linguistic-Enhanced Transformer with CTC Embedding for Speech Recognition25 Oct 2022 0 repositories listed
-
Rich Knowledge Sources Bring Complex Knowledge Conflicts: Recalibrating Models to Reflect Conflicting Evidence25 Oct 2022 0 repositories listed
-
Same Pre-training Loss, Better Downstream: Implicit Bias Matters for Language Models25 Oct 2022 0 repositories listed
-
Towards Better Few-Shot and Finetuning Performance with Forgetful Causal Language Models24 Oct 2022 0 repositories listed
-
A BERT-based Deep Learning Approach for Reputation Analysis in Social Media23 Oct 2022 0 repositories listed
-
Discriminative Language Model as Semantic Consistency Scorer for Prompt-based Few-Shot Text Classification23 Oct 2022 0 repositories listed
-
Do Language Models Understand Measurements?23 Oct 2022 0 repositories listed
-
Hard Gate Knowledge Distillation -- Leverage Calibration for Robust and Reliable Language Model22 Oct 2022 0 repositories listed
-
LMPriors: Pre-Trained Language Models as Task-Specific Priors22 Oct 2022 0 repositories listed
-
P³LM: Probabilistically Permuted Prophet Language Modeling for Generative Pre-Training22 Oct 2022 0 repositories listed
-
PENTATRON: PErsonalized coNText-Aware Transformer for Retrieval-based cOnversational uNderstanding22 Oct 2022 0 repositories listed