Browse State-of-the-Art › Language Modeling › Papers, page 108
Language Modeling
Papers archive 2025-07-28
archive papers tagged: 14,182 · with a code link: 5,620 · where Syntology ran a sample: 1,894 (1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,894 of 14,182 tagged: 1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument)
Page 108 of 142: papers 10,701 to 10,800 of 14,182, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
From Retrieval to Generation: Efficient and Effective Entity Set Expansion7 Apr 2023 0 repositories listed
-
TemPL: A Novel Deep Learning Model for Zero-Shot Prediction of Protein Stability and Activity Based on Temperature-Guided Language Modeling7 Apr 2023 0 repositories listed
-
Revolutionizing Single Cell Analysis: The Power of Large Language Models for Cell Type Annotation5 Apr 2023 0 repositories listed
-
Towards Self-Explainability of Deep Neural Networks with Heatmap Captioning and Large-Language Models5 Apr 2023 0 repositories listed
-
Dialogue-Contextualized Re-ranking for Medical History-Taking4 Apr 2023 0 repositories listed
-
Is ChatGPT a Highly Fluent Grammatical Error Correction System? A Comprehensive Evaluation4 Apr 2023 0 repositories listed
-
Using Language Models For Knowledge Acquisition in Natural Language Reasoning Problems4 Apr 2023 0 repositories listed
-
Multi-Modal Perceiver Language Model for Outcome Prediction in Emergency Department3 Apr 2023 0 repositories listed
-
Demonstration of InsightPilot: An LLM-Empowered Automated Data Exploration System2 Apr 2023 0 repositories listed
-
A Measurement-Based Quantum-Like Language Model for Text Matching1 Apr 2023 0 repositories listed
-
Network Visualization of ChatGPT Research: a study based on term and keyword co-occurrence network analysis1 Apr 2023 0 repositories listed
-
Quick Dense Retrievers Consume KALE: Post Training Kullback Leibler Alignment of Embeddings for Asymmetrical dual encoders31 Mar 2023 0 repositories listed
-
A BERT-based Unsupervised Grammatical Error Correction Framework30 Mar 2023 0 repositories listed
-
The Nordic Pile: A 1.2TB Nordic Dataset for Language Modeling30 Mar 2023 0 repositories listed
-
Advances in apparent conceptual physics reasoning in GPT-429 Mar 2023 0 repositories listed
-
Joint unsupervised and supervised learning for context-aware language identification29 Mar 2023 0 repositories listed
-
Mask-free OVIS: Open-Vocabulary Instance Segmentation without Manual Mask Annotations29 Mar 2023 0 repositories listed
-
ProtFIM: Fill-in-Middle Protein Sequence Design via Protein Language Models29 Mar 2023 0 repositories listed
-
Planning with Sequence Models through Iterative Energy Minimization28 Mar 2023 0 repositories listed
-
Structured Video-Language Modeling with Temporal Grouping and Spatial Grounding28 Mar 2023 0 repositories listed
-
Cross-utterance ASR Rescoring with Graph-based Label Propagation27 Mar 2023 0 repositories listed
-
Linguistically Informed ChatGPT Prompts to Enhance Japanese-Chinese Machine Translation: A Case Study on Attributive Clauses27 Mar 2023 0 repositories listed
-
LMCanvas: Object-Oriented Interaction to Personalize Large Language Model-Powered Writing Environments27 Mar 2023 0 repositories listed
-
Typhoon: Towards an Effective Task-Specific Masking Strategy for Pre-trained Language Models27 Mar 2023 0 repositories listed
-
Sem4SAP: Synonymous Expression Mining From Open Knowledge Graph For Language Model Synonym-Aware Pretraining25 Mar 2023 0 repositories listed
-
Attention-based Speech Enhancement Using Human Quality Perception Modelling23 Mar 2023 0 repositories listed
-
ChatGPT for Shaping the Future of Dentistry: The Potential of Multi-Modal Large Language Model23 Mar 2023 0 repositories listed
-
SPeC: A Soft Prompt-Based Calibration on Performance Variability of Large Language Model in Clinical Notes Summarization23 Mar 2023 0 repositories listed
-
Can we trust the evaluation on ChatGPT?22 Mar 2023 0 repositories listed
-
Frozen Language Model Helps ECG Zero-Shot Learning22 Mar 2023 0 repositories listed
-
Salient Span Masking for Temporal Understanding22 Mar 2023 0 repositories listed
-
HOP+: History-enhanced and Order-aware Pre-training for Vision-and-Language Navigation20 Mar 2023 0 repositories listed
-
Large Language Models and Simple, Stupid Bugs20 Mar 2023 0 repositories listed
-
Multimodal Shannon Game with Images20 Mar 2023 0 repositories listed
-
On-the-fly Text Retrieval for End-to-End ASR Adaptation20 Mar 2023 0 repositories listed
-
PanGu-Σ: Towards Trillion Parameter Language Model with Sparse Heterogeneous Computing20 Mar 2023 0 repositories listed
-
Label Name is Mantra: Unifying Point Cloud Segmentation across Heterogeneous Datasets19 Mar 2023 0 repositories listed
-
CCPL: Cross-modal Contrastive Protein Learning19 Mar 2023 0 repositories listed
-
DORIC : Domain Robust Fine-Tuning for Open Intent Clustering through Dependency Parsing17 Mar 2023 0 repositories listed
-
Towards the Scalable Evaluation of Cooperativeness in Language Models16 Mar 2023 0 repositories listed
-
ChatGPT or Grammarly? Evaluating ChatGPT on Grammatical Error Correction Benchmark15 Mar 2023 0 repositories listed
-
Contextualized Medication Information Extraction Using Transformer-based Deep Learning Architectures14 Mar 2023 0 repositories listed
-
Do Transformers Parse while Predicting the Masked Word?14 Mar 2023 0 repositories listed
-
Finding the Needle in a Haystack: Unsupervised Rationale Extraction from Long Text Classifiers14 Mar 2023 0 repositories listed
-
Generating multiple-choice questions for medical question answering with distractors and cue-masking13 Mar 2023 0 repositories listed
-
ODIN: On-demand Data Formulation to Mitigate Dataset Lock-in13 Mar 2023 0 repositories listed
-
Learning Combinatorial Prompts for Universal Controllable Image Captioning11 Mar 2023 0 repositories listed
-
Algorithmic Ghost in the Research Shell: Large Language Models and Academic Knowledge Creation in Management Research10 Mar 2023 0 repositories listed
-
An Overview on Language Models: Recent Developments and Outlook10 Mar 2023 0 repositories listed
-
Towards MoE Deployment: Mitigating Inefficiencies in Mixture-of-Expert (MoE) Inference10 Mar 2023 0 repositories listed
-
9 Mar 2023 0 repositories listed
-
Knowledge-augmented Few-shot Visual Relation Detection9 Mar 2023 0 repositories listed
-
Refined Vision-Language Modeling for Fine-grained Multi-modal Pre-training9 Mar 2023 0 repositories listed
-
Weakly-Supervised HOI Detection from Interaction Labels Only and Language/Vision-Language Priors9 Mar 2023 0 repositories listed
-
Extending the Pre-Training of BLOOM for Improved Support of Traditional Chinese: Models, Methods and Results8 Mar 2023 0 repositories listed
-
Magnushammer: A Transformer-Based Approach to Premise Selection8 Mar 2023 0 repositories listed
-
ChatGPT: Beginning of an End of Manual Linguistic Data Annotation? Use Case of Automatic Genre Identification7 Mar 2023 0 repositories listed
-
Making a Computational Attorney7 Mar 2023 0 repositories listed
-
ChatGPT is on the Horizon: Could a Large Language Model be Suitable for Intelligent Traffic Safety Research and Applications?6 Mar 2023 0 repositories listed
-
Data Portraits: Recording Foundation Model Training Data6 Mar 2023 0 repositories listed
-
FoundationTTS: Text-to-Speech for ASR Customization with Generative Language Model6 Mar 2023 0 repositories listed
-
Model-Agnostic Meta-Learning for Natural Language Understanding Tasks in Finance6 Mar 2023 0 repositories listed
-
Spelling convention sensitivity in neural language models6 Mar 2023 0 repositories listed
-
Could a Large Language Model be Conscious?4 Mar 2023 0 repositories listed
-
End-to-End Speech Recognition: A Survey3 Mar 2023 0 repositories listed
-
RePreM: Representation Pre-training with Masked Model for Reinforcement Learning3 Mar 2023 0 repositories listed
-
Will Affective Computing Emerge from Foundation Models and General AI? A First Evaluation on ChatGPT3 Mar 2023 0 repositories listed
-
BenchDirect: A Directed Language Model for Compiler Benchmarks2 Mar 2023 0 repositories listed
-
How will Language Modelers like ChatGPT Affect Occupations and Industries?2 Mar 2023 0 repositories listed
-
Semiparametric Language Models Are Scalable Continual Learners2 Mar 2023 0 repositories listed
-
Almanac: Retrieval-Augmented Language Models for Clinical Medicine1 Mar 2023 0 repositories listed
-
Domain-adapted large language models for classifying nuclear medicine reports1 Mar 2023 0 repositories listed
-
Grounded Decoding: Guiding Text Generation with Grounded Models for Embodied Agents1 Mar 2023 0 repositories listed
-
N-best T5: Robust ASR Error Correction using Multiple Input Hypotheses and Constrained Decoding Space1 Mar 2023 0 repositories listed
-
1 Mar 2023 0 repositories listed
-
Efficient Masked Autoencoders with Self-Consistency28 Feb 2023 0 repositories listed
-
Weighted Sampling for Masked Language Modeling28 Feb 2023 0 repositories listed
-
Duration-aware pause insertion using pre-trained language model for multi-speaker text-to-speech27 Feb 2023 0 repositories listed
-
Leveraging Large Language Model and Story-Based Gamification in Intelligent Tutoring System to Scaffold Introductory Programming Courses: A Design-Based Research Study25 Feb 2023 0 repositories listed
-
Topic-Selective Graph Network for Topic-Focused Summarization25 Feb 2023 0 repositories listed
-
Toward Fairness in Text Generation via Mutual Information Minimization based on Importance Sampling25 Feb 2023 0 repositories listed
-
Factual Consistency Oriented Speech Recognition24 Feb 2023 0 repositories listed
-
Generative Sentiment Transfer via Adaptive Masking23 Feb 2023 0 repositories listed
-
On the Generalization Ability of Retrieval-Enhanced Transformers23 Feb 2023 0 repositories listed
-
23 Feb 2023 0 repositories listed
-
BadGPT: Exploring Security Vulnerabilities of ChatGPT via Backdoor Attacks to InstructGPT21 Feb 2023 0 repositories listed
-
kNN-Adapter: Efficient Domain Adaptation for Black-Box Language Models21 Feb 2023 0 repositories listed
-
Bag of Tricks for Effective Language Model Pretraining and Downstream Adaptation: A Case Study on GLUE18 Feb 2023 0 repositories listed
-
Multiperiodic Processes: Ergodic Sources with a Sublinear Entropy17 Feb 2023 0 repositories listed
-
Entry Separation using a Mixed Visual and Textual Language Model: Application to 19th century French Trade Directories17 Feb 2023 0 repositories listed
-
GPT4MIA: Utilizing Generative Pre-trained Transformer (GPT-3) as A Plug-and-Play Transductive Model for Medical Image Analysis17 Feb 2023 0 repositories listed
-
Massively Multilingual Shallow Fusion with Large Language Models17 Feb 2023 0 repositories listed
-
Privately Customizing Prefinetuning to Better Match User Data in Federated Learning17 Feb 2023 0 repositories listed
-
Prompting Large Language Models With the Socratic Method17 Feb 2023 0 repositories listed
-
Adaptable End-to-End ASR Models using Replaceable Internal LMs and Residual Softmax16 Feb 2023 0 repositories listed
-
Bridge the Gap between Language models and Tabular Understanding16 Feb 2023 0 repositories listed
-
JEIT: Joint End-to-End Model and Internal Language Model Training for Speech Recognition16 Feb 2023 0 repositories listed
-
Learning to Initialize: Can Meta Learning Improve Cross-task Generalization in Prompt Tuning?16 Feb 2023 0 repositories listed
-
Role of Bias Terms in Dot-Product Attention16 Feb 2023 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.