Browse State-of-the-Art › Language Modeling › Papers, page 100
Language Modeling
Papers archive 2025-07-28
archive papers tagged: 14,182 · with a code link: 5,620 · where Syntology ran a sample: 1,894 (1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,894 of 14,182 tagged: 1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument)
Page 100 of 142: papers 9,901 to 10,000 of 14,182, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Frontier Language Models are not Robust to Adversarial Arithmetic, or "What do I need to say so you agree 2+2=5?8 Nov 2023 0 repositories listed
-
AI for All: Operationalising Diversity and Inclusion Requirements for AI Systems7 Nov 2023 0 repositories listed
-
Evaluating the Effectiveness of Retrieval-Augmented Large Language Models in Scientific Document Reasoning7 Nov 2023 0 repositories listed
-
Formal Aspects of Language Modeling7 Nov 2023 0 repositories listed
-
OLaLa: Ontology Matching with Large Language Models7 Nov 2023 0 repositories listed
-
DAIL: Data Augmentation for In-Context Learning via Self-Paraphrase6 Nov 2023 0 repositories listed
-
Leveraging High-Level Synthesis and Large Language Models to Generate, Simulate, and Deploy a Uniform Random Number Generator Hardware Design6 Nov 2023 0 repositories listed
-
ProPath: Disease-Specific Protein Language Model for Variant Pathogenicity6 Nov 2023 0 repositories listed
-
Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation6 Nov 2023 0 repositories listed
-
CIRCLE: Multi-Turn Query Clarifications with Reinforcement Learning5 Nov 2023 0 repositories listed
-
Large language models implicitly learn to straighten neural sentence trajectories to construct a predictive representation of natural language5 Nov 2023 0 repositories listed
-
Can Chat GPT solve a Linguistics Exam?4 Nov 2023 0 repositories listed
-
Understanding the Natural Language of DNA using Encoder-Decoder Foundation Models with Byte-level Precision4 Nov 2023 0 repositories listed
-
COSMIC: Data Efficient Instruction-tuning For Speech In-Context Learning3 Nov 2023 0 repositories listed
-
Data-Free Distillation of Language Model by Text-to-Text Transfer3 Nov 2023 0 repositories listed
-
Supermind Ideator: Exploring generative AI to support creative problem-solving3 Nov 2023 0 repositories listed
-
Too Much Information: Keeping Training Simple for BabyLMs3 Nov 2023 0 repositories listed
-
Continual Learning Under Language Shift2 Nov 2023 0 repositories listed
-
Expressive TTS Driven by Natural Language Prompts Using Few Human Annotations2 Nov 2023 0 repositories listed
-
FlashDecoding++: Faster Large Language Model Inference on GPUs2 Nov 2023 0 repositories listed
-
Predicting Question-Answering Performance of Large Language Models through Semantic Consistency2 Nov 2023 0 repositories listed
-
Recommendations by Concise User Profiles from Review Text2 Nov 2023 0 repositories listed
-
Self-Influence Guided Data Reweighting for Language Model Pre-training2 Nov 2023 0 repositories listed
-
An Improved Transformer-based Model for Detecting Phishing, Spam, and Ham: A Large Language Model Approach1 Nov 2023 0 repositories listed
-
Attention Alignment and Flexible Positional Embeddings Improve Transformer Length Extrapolation1 Nov 2023 0 repositories listed
-
CLIP-AD: A Language-Guided Staged Dual-Path Model for Zero-shot Anomaly Detection1 Nov 2023 0 repositories listed
-
Form follows Function: Text-to-Text Conditional Graph Generation based on Functional Requirements1 Nov 2023 0 repositories listed
-
Modeling subjectivity (by Mimicking Annotator Annotation) in toxic comment identification across diverse communities1 Nov 2023 0 repositories listed
-
ZEETAD: Adapting Pretrained Vision-Language Model for Zero-Shot End-to-End Temporal Action Detection1 Nov 2023 0 repositories listed
-
BERTwich: Extending BERT's Capabilities to Model Dialectal and Noisy Text31 Oct 2023 0 repositories listed
-
Enhancing the Spatial Awareness Capability of Multi-Modal Large Language Model31 Oct 2023 0 repositories listed
-
FA Team at the NTCIR-17 UFO Task31 Oct 2023 0 repositories listed
-
Filter bubbles and affective polarization in user-personalized large language model outputs31 Oct 2023 0 repositories listed
-
Interactive Multi-fidelity Learning for Cost-effective Adaptation of Language Model with Sparse Human Supervision31 Oct 2023 0 repositories listed
-
Longer Fixations, More Computation: Gaze-Guided Recurrent Neural Networks31 Oct 2023 0 repositories listed
-
A Multi-Modal Foundation Model to Assist People with Blindness and Low Vision in Environmental Interaction31 Oct 2023 0 repositories listed
-
Adapter Pruning using Tropical Characterization30 Oct 2023 0 repositories listed
-
EHRTutor: Enhancing Patient Understanding of Discharge Instructions30 Oct 2023 0 repositories listed
-
Generative retrieval-augmented ontologic graph and multi-agent strategies for interpretive large language model-based materials design30 Oct 2023 0 repositories listed
-
Improving Input-label Mapping with Demonstration Replay for In-context Learning30 Oct 2023 0 repositories listed
-
30 Oct 2023 0 repositories listed
-
Leveraging Language Models to Detect Greenwashing30 Oct 2023 0 repositories listed
-
30 Oct 2023 0 repositories listed Syntology 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Musical Form Generation30 Oct 2023 0 repositories listed
-
The Impact of Depth on Compositional Generalization in Transformer Language Models30 Oct 2023 0 repositories listed
-
Robustifying Language Models with Test-Time Adaptation29 Oct 2023 0 repositories listed
-
TeacherLM: Teaching to Fish Rather Than Giving the Fish, Language Modeling Likewise29 Oct 2023 0 repositories listed
-
Reboost Large Language Model-based Text-to-SQL, Text-to-Python, and Text-to-Function -- with Real Applications in Traffic Domain28 Oct 2023 0 repositories listed
-
Robust NL-to-Cypher Translation for KBQA: Harnessing Large Language Model with Chain of Prompts28 Oct 2023 0 repositories listed
-
Generative AI for Software Metadata: Overview of the Information Retrieval in Software Engineering Track at FIRE 202327 Oct 2023 0 repositories listed
-
Interactive Robot Learning from Verbal Correction26 Oct 2023 0 repositories listed
-
Large Language Models as Generalizable Policies for Embodied Tasks26 Oct 2023 0 repositories listed
-
BOOST: Harnessing Black-Box Control to Boost Commonsense in LMs' Generation25 Oct 2023 0 repositories listed
-
Controlled Decoding from Language Models25 Oct 2023 0 repositories listed
-
Faithful Path Language Modeling for Explainable Recommendation over Knowledge Graph25 Oct 2023 0 repositories listed
-
FedTherapist: Mental Health Monitoring with User-Generated Linguistic Expressions on Smartphones via Federated Learning25 Oct 2023 0 repositories listed
-
General Point Model with Autoencoding and Autoregressive25 Oct 2023 0 repositories listed
-
math-PVS: A Large Language Model Framework to Map Scientific Publications to PVS Theories25 Oct 2023 0 repositories listed
-
Multiple Key-value Strategy in Recommendation Systems Incorporating Large Language Model25 Oct 2023 0 repositories listed
-
RCAgent: Cloud Root Cause Analysis by Autonomous Agents with Tool-Augmented Large Language Models25 Oct 2023 0 repositories listed
-
Subspace Chronicles: How Linguistic Information Emerges, Shifts and Interacts during Language Model Training25 Oct 2023 0 repositories listed
-
Transformer-based Live Update Generation for Soccer Matches from Microblog Posts25 Oct 2023 0 repositories listed
-
URL-BERT: Training Webpage Representations via Social Media Engagements25 Oct 2023 0 repositories listed
-
A Language Model with Limited Memory Capacity Captures Interference in Human Sentence Processing24 Oct 2023 0 repositories listed
-
BLP-2023 Task 2: Sentiment Analysis24 Oct 2023 0 repositories listed
-
DeSIQ: Towards an Unbiased, Challenging Benchmark for Social Intelligence Understanding24 Oct 2023 0 repositories listed
-
E-Sparse: Boosting the Large Language Model Inference through Entropy-based N:M Sparsity24 Oct 2023 0 repositories listed
-
Facilitating Self-Guided Mental Health Interventions Through Human-Language Model Interaction: A Case Study of Cognitive Restructuring24 Oct 2023 0 repositories listed
-
FLTrojan: Privacy Leakage Attacks against Federated Language Models Through Selective Weight Tampering24 Oct 2023 0 repositories listed
-
MindLLM: Pre-training Lightweight Large Language Model from Scratch, Evaluations and Domain Applications24 Oct 2023 0 repositories listed
-
Prevalence and prevention of large language model use in crowd work24 Oct 2023 0 repositories listed
-
PromptInfuser: How Tightly Coupling AI and UI Design Impacts Designers' Workflows24 Oct 2023 0 repositories listed
-
Retrieval-based Knowledge Transfer: An Effective Approach for Extreme Large Language Model Compression24 Oct 2023 0 repositories listed
-
Rosetta Stone at KSAA-RD Shared Task: A Hop From Language Modeling To Word--Definition Alignment24 Oct 2023 0 repositories listed
-
TCRA-LLM: Token Compression Retrieval Augmented Large Language Model for Inference Cost Reduction24 Oct 2023 0 repositories listed
-
Unnatural language processing: How do language models handle machine-generated prompts?24 Oct 2023 0 repositories listed
-
WebWISE: Web Interface Control and Sequential Exploration with Large Language Models24 Oct 2023 0 repositories listed
-
Branch-Solve-Merge Improves Large Language Model Evaluation and Generation23 Oct 2023 0 repositories listed
-
Counting the Bugs in ChatGPT's Wugs: A Multilingual Investigation into the Morphological Capabilities of a Large Language Model23 Oct 2023 0 repositories listed
-
Generative Pre-trained Transformer for Vietnamese Community-based COVID-19 Question Answering23 Oct 2023 0 repositories listed
-
Health Disparities through Generative AI Models: A Comparison Study Using A Domain Specific large language model23 Oct 2023 0 repositories listed
-
Irreducible Curriculum for Language Model Pretraining23 Oct 2023 0 repositories listed
-
Large Search Model: Redefining Search Stack in the Era of LLMs23 Oct 2023 0 repositories listed
-
Understanding the Inner Workings of Language Models Through Representation Dissimilarity23 Oct 2023 0 repositories listed
-
Why LLMs Hallucinate, and How to Get (Evidential) Closure: Perceptual, Intensional, and Extensional Learning for Faithful Natural Language Generation23 Oct 2023 0 repositories listed
-
Boosting Unsupervised Machine Translation with Pseudo-Parallel Data22 Oct 2023 0 repositories listed
-
Customising General Large Language Models for Specialised Emotion Recognition Tasks22 Oct 2023 0 repositories listed
-
One Model for All: Large Language Models are Domain-Agnostic Recommendation Systems22 Oct 2023 0 repositories listed
-
Which Prompts Make The Difference? Data Prioritization For Efficient Human LLM Evaluation22 Oct 2023 0 repositories listed
-
Learning Reward for Physical Skills using Large Language Model21 Oct 2023 0 repositories listed
-
MedEval: A Multi-Level, Multi-Task, and Multi-Domain Medical Benchmark for Language Model Evaluation21 Oct 2023 0 repositories listed Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Sentiment Analysis Across Multiple African Languages: A Current Benchmark21 Oct 2023 0 repositories listed
-
Ask Language Model to Clean Your Noisy Translation Data20 Oct 2023 0 repositories listed
-
Cache & Distil: Optimising API Calls to Large Language Models20 Oct 2023 0 repositories listed
-
Enhancing Zero-Shot Crypto Sentiment with Fine-tuned Language Model and Prompt Engineering20 Oct 2023 0 repositories listed
-
GenDistiller: Distilling Pre-trained Language Models based on Generative Models20 Oct 2023 0 repositories listed
-
The Past, Present, and Future of Typological Databases in NLP20 Oct 2023 0 repositories listed
-
Thoroughly Modeling Multi-domain Pre-trained Recommendation as Language20 Oct 2023 0 repositories listed
-
WordArt Designer: User-Driven Artistic Typography Synthesis using Large Language Models20 Oct 2023 0 repositories listed
-
Zero-Shot Sharpness-Aware Quantization for Pre-trained Language Models20 Oct 2023 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.