Browse State-of-the-Art › Language Modelling › Papers, page 121
Language Modelling
Papers archive 2025-07-28
archive papers tagged: 17,610 · with a code link: 7,012 · where Syntology ran a sample: 2,428 (2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,428 of 17,610 tagged: 2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument)
Page 121 of 177: papers 12,001 to 12,100 of 17,610, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
math-PVS: A Large Language Model Framework to Map Scientific Publications to PVS Theories25 Oct 2023 0 repositories listed
-
Multiple Key-value Strategy in Recommendation Systems Incorporating Large Language Model25 Oct 2023 0 repositories listed
-
RCAgent: Cloud Root Cause Analysis by Autonomous Agents with Tool-Augmented Large Language Models25 Oct 2023 0 repositories listed
-
Subspace Chronicles: How Linguistic Information Emerges, Shifts and Interacts during Language Model Training25 Oct 2023 0 repositories listed
-
Optimal Inflationary Potentials25 Oct 2023 0 repositories listed
-
Transformer-based Live Update Generation for Soccer Matches from Microblog Posts25 Oct 2023 0 repositories listed
-
URL-BERT: Training Webpage Representations via Social Media Engagements25 Oct 2023 0 repositories listed
-
Using GPT-4 to Augment Unbalanced Data for Automatic Scoring25 Oct 2023 0 repositories listed
-
A Language Model with Limited Memory Capacity Captures Interference in Human Sentence Processing24 Oct 2023 0 repositories listed
-
BLP-2023 Task 2: Sentiment Analysis24 Oct 2023 0 repositories listed
-
ConstitutionMaker: Interactively Critiquing Large Language Models by Converting Feedback into Principles24 Oct 2023 0 repositories listed
-
DeSIQ: Towards an Unbiased, Challenging Benchmark for Social Intelligence Understanding24 Oct 2023 0 repositories listed
-
E-Sparse: Boosting the Large Language Model Inference through Entropy-based N:M Sparsity24 Oct 2023 0 repositories listed
-
Facilitating Self-Guided Mental Health Interventions Through Human-Language Model Interaction: A Case Study of Cognitive Restructuring24 Oct 2023 0 repositories listed
-
FLTrojan: Privacy Leakage Attacks against Federated Language Models Through Selective Weight Tampering24 Oct 2023 0 repositories listed
-
Leveraging Large Language Models for Enhanced Product Descriptions in eCommerce24 Oct 2023 0 repositories listed
-
MindLLM: Pre-training Lightweight Large Language Model from Scratch, Evaluations and Domain Applications24 Oct 2023 0 repositories listed
-
Prevalence and prevention of large language model use in crowd work24 Oct 2023 0 repositories listed
-
PromptInfuser: How Tightly Coupling AI and UI Design Impacts Designers' Workflows24 Oct 2023 0 repositories listed
-
Retrieval-based Knowledge Transfer: An Effective Approach for Extreme Large Language Model Compression24 Oct 2023 0 repositories listed
-
Rosetta Stone at KSAA-RD Shared Task: A Hop From Language Modeling To Word--Definition Alignment24 Oct 2023 0 repositories listed
-
Self-Guard: Empower the LLM to Safeguard Itself24 Oct 2023 0 repositories listed
-
TCRA-LLM: Token Compression Retrieval Augmented Large Language Model for Inference Cost Reduction24 Oct 2023 0 repositories listed
-
Unnatural language processing: How do language models handle machine-generated prompts?24 Oct 2023 0 repositories listed
-
WebWISE: Web Interface Control and Sequential Exploration with Large Language Models24 Oct 2023 0 repositories listed
-
Branch-Solve-Merge Improves Large Language Model Evaluation and Generation23 Oct 2023 0 repositories listed
-
Counting the Bugs in ChatGPT's Wugs: A Multilingual Investigation into the Morphological Capabilities of a Large Language Model23 Oct 2023 0 repositories listed
-
DoGE: Domain Reweighting with Generalization Estimation23 Oct 2023 0 repositories listed
-
Generative Pre-trained Transformer for Vietnamese Community-based COVID-19 Question Answering23 Oct 2023 0 repositories listed
-
Health Disparities through Generative AI Models: A Comparison Study Using A Domain Specific large language model23 Oct 2023 0 repositories listed
-
Irreducible Curriculum for Language Model Pretraining23 Oct 2023 0 repositories listed
-
Large Search Model: Redefining Search Stack in the Era of LLMs23 Oct 2023 0 repositories listed
-
SpecTr: Fast Speculative Decoding via Optimal Transport23 Oct 2023 0 repositories listed
-
SuperTweetEval: A Challenging, Unified and Heterogeneous Benchmark for Social Media NLP Research23 Oct 2023 0 repositories listed
-
Understanding the Inner Workings of Language Models Through Representation Dissimilarity23 Oct 2023 0 repositories listed
-
Why LLMs Hallucinate, and How to Get (Evidential) Closure: Perceptual, Intensional, and Extensional Learning for Faithful Natural Language Generation23 Oct 2023 0 repositories listed
-
Boosting Unsupervised Machine Translation with Pseudo-Parallel Data22 Oct 2023 0 repositories listed
-
Customising General Large Language Models for Specialised Emotion Recognition Tasks22 Oct 2023 0 repositories listed
-
MoPe: Model Perturbation-based Privacy Attacks on Language Models22 Oct 2023 0 repositories listed
-
Neural Text Sanitization with Privacy Risk Indicators: An Empirical Analysis22 Oct 2023 0 repositories listed
-
One Model for All: Large Language Models are Domain-Agnostic Recommendation Systems22 Oct 2023 0 repositories listed
-
Which Prompts Make The Difference? Data Prioritization For Efficient Human LLM Evaluation22 Oct 2023 0 repositories listed
-
Learning Reward for Physical Skills using Large Language Model21 Oct 2023 0 repositories listed
-
MedEval: A Multi-Level, Multi-Task, and Multi-Domain Medical Benchmark for Language Model Evaluation21 Oct 2023 0 repositories listed Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Sentiment Analysis Across Multiple African Languages: A Current Benchmark21 Oct 2023 0 repositories listed
-
Ask Language Model to Clean Your Noisy Translation Data20 Oct 2023 0 repositories listed
-
Cache & Distil: Optimising API Calls to Large Language Models20 Oct 2023 0 repositories listed
-
Conversation Chronicles: Towards Diverse Temporal and Relational Dynamics in Multi-Session Conversations20 Oct 2023 0 repositories listed
-
Enhancing Zero-Shot Crypto Sentiment with Fine-tuned Language Model and Prompt Engineering20 Oct 2023 0 repositories listed
-
FABULA: Intelligence Report Generation Using Retrieval-Augmented Narrative Construction20 Oct 2023 0 repositories listed
-
GenDistiller: Distilling Pre-trained Language Models based on Generative Models20 Oct 2023 0 repositories listed
-
The Past, Present, and Future of Typological Databases in NLP20 Oct 2023 0 repositories listed
-
Thoroughly Modeling Multi-domain Pre-trained Recommendation as Language20 Oct 2023 0 repositories listed
-
WordArt Designer: User-Driven Artistic Typography Synthesis using Large Language Models20 Oct 2023 0 repositories listed
-
Zero-Shot Sharpness-Aware Quantization for Pre-trained Language Models20 Oct 2023 0 repositories listed
-
A Systematic Study of Performance Disparities in Multilingual Task-Oriented Dialogue Systems19 Oct 2023 0 repositories listed
-
Exploring In-Context Learning of Textless Speech Language Model for Speech Classification Tasks19 Oct 2023 0 repositories listed
-
ReEval: Automatic Hallucination Evaluation for Retrieval-Augmented Large Language Models via Transferable Adversarial Attacks19 Oct 2023 0 repositories listed
-
Data Augmentations for Improved (Large) Language Model Generalization19 Oct 2023 0 repositories listed
-
CLAIR: Evaluating Image Captions with Large Language Models19 Oct 2023 0 repositories listed
-
Efficient Long-Range Transformers: You Need to Attend More, but Not Necessarily at Every Layer19 Oct 2023 0 repositories listed
-
Is ChatGPT a Financial Expert? Evaluating Language Models on Financial Natural Language Processing19 Oct 2023 0 repositories listed
-
Label-Aware Automatic Verbalizer for Few-Shot Text Classification19 Oct 2023 0 repositories listed
-
LASER: Linear Compression in Wireless Distributed Optimization19 Oct 2023 0 repositories listed
-
Lost in Translation: When GPT-4V(ision) Can't See Eye to Eye with Text. A Vision-Language-Consistency Analysis of VLLMs and Beyond19 Oct 2023 0 repositories listed
-
Named Entity Recognition for Monitoring Plant Health Threats in Tweets: a ChouBERT Approach19 Oct 2023 0 repositories listed
-
Document-Level Language Models for Machine Translation18 Oct 2023 0 repositories listed
-
Pseudointelligence: A Unifying Framework for Language Model Evaluation18 Oct 2023 0 repositories listed
-
Solving the multiplication problem of a large language model system using a graph-based method18 Oct 2023 0 repositories listed
-
ChapGTP, ILLC's Attempt at Raising a BabyLM: Improving Data Efficiency by Automatic Task Formation17 Oct 2023 0 repositories listed
-
Correction Focused Language Model Training for Speech Recognition17 Oct 2023 0 repositories listed
-
17 Oct 2023 0 repositories listed
-
Generative error correction for code-switching speech recognition using large language models17 Oct 2023 0 repositories listed
-
Iterative Shallow Fusion of Backward Language Model for End-to-End Speech Recognition17 Oct 2023 0 repositories listed
-
Large Language Model Prediction Capabilities: Evidence from a Real-World Forecasting Tournament17 Oct 2023 0 repositories listed
-
Leveraging Large Language Model for Automatic Evolving of Industrial Data-Centric R&D Cycle17 Oct 2023 0 repositories listed
-
Multi-stage Large Language Model Correction for Speech Recognition17 Oct 2023 0 repositories listed
-
Revealing the Unwritten: Visual Investigation of Beam Search Trees to Address Language Model Prompting Challenges17 Oct 2023 0 repositories listed
-
Utilising a Large Language Model to Annotate Subject Metadata: A Case Study in an Australian National Research Data Catalogue17 Oct 2023 0 repositories listed
-
Bootstrap Your Own Skills: Learning to Solve New Tasks with Large Language Model Guidance16 Oct 2023 0 repositories listed
-
Contextual Data Augmentation for Task-Oriented Dialog Systems16 Oct 2023 0 repositories listed
-
EfficientOCR: An Extensible, Open-Source Package for Efficiently Digitizing World Knowledge16 Oct 2023 0 repositories listed
-
Fine-tuning ChatGPT for Automatic Scoring16 Oct 2023 0 repositories listed
-
ForceGen: End-to-end de novo protein generation based on nonlinear mechanical unfolding responses using a protein language diffusion model16 Oct 2023 0 repositories listed
-
Improving Large Language Model Fine-tuning for Solving Math Problems16 Oct 2023 0 repositories listed
-
Interactive Task Planning with Language Models16 Oct 2023 0 repositories listed
-
DavIR: Data Selection via Implicit Reward for Large Language Models16 Oct 2023 0 repositories listed
-
MechGPT, a language-based strategy for mechanics and materials modeling that connects knowledge across scales, disciplines and modalities16 Oct 2023 0 repositories listed
-
Navigation with Large Language Models: Semantic Guesswork as a Heuristic for Planning16 Oct 2023 0 repositories listed
-
Use of probabilistic phrases in a coordination game: human versus GPT-416 Oct 2023 0 repositories listed
-
ACES: Generating Diverse Programming Puzzles with with Autotelic Generative Models15 Oct 2023 0 repositories listed
-
Beyond Segmentation: Road Network Generation with Multi-Modal LLMs15 Oct 2023 0 repositories listed
-
Empirical study of pretrained multilingual language models for zero-shot cross-lingual knowledge transfer in generation15 Oct 2023 0 repositories listed
-
Farzi Data: Autoregressive Data Distillation15 Oct 2023 0 repositories listed
-
Large Language Model-Aware In-Context Learning for Code Generation15 Oct 2023 0 repositories listed
-
Large Vocabulary Spontaneous Speech Recognition for Tigrigna15 Oct 2023 0 repositories listed
-
Reformulating NLP tasks to Capture Longitudinal Manifestation of Language Disorders in People with Dementia15 Oct 2023 0 repositories listed
-
A study of the impact of generative AI-based data augmentation on software metadata classification14 Oct 2023 0 repositories listed
-
Chatbot-supported Thesis Writing: An Autoethnographic Report14 Oct 2023 0 repositories listed
-
Enhancing Binary Code Comment Quality Classification: Integrating Generative AI for Improved Accuracy14 Oct 2023 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.