Browse State-of-the-Art › Language Modelling › Papers, page 90
Language Modelling
Papers archive 2025-07-28
archive papers tagged: 17,610 · with a code link: 7,012 · where Syntology ran a sample: 2,428 (2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (2,428 of 17,610 tagged: 2,027 with a run with no instrument failure, 401 where every run was a failure of Syntology's instrument)
Page 90 of 177: papers 8,901 to 9,000 of 17,610, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
SSSD: Simply-Scalable Speculative Decoding8 Nov 2024 0 repositories listed
-
The Empirical Impact of Data Sanitization on Language Models8 Nov 2024 0 repositories listed
-
Towards Multi-Modal Mastery: A 4.5B Parameter Truly Multi-Modal Small Language Model8 Nov 2024 0 repositories listed
-
Unmasking the Shadows: Pinpoint the Implementations of Anti-Dynamic Analysis Techniques in Malware Using LLM8 Nov 2024 0 repositories listed
-
A Reinforcement Learning-Based Automatic Video Editing Method Using Pre-trained Vision-Language Model7 Nov 2024 0 repositories listed
-
Benchmarking Large Language Models with Integer Sequence Generation Tasks7 Nov 2024 0 repositories listed
-
CUIfy the XR: An Open-Source Package to Embed LLM-powered Conversational Agents in XR7 Nov 2024 0 repositories listed
-
Scaling Laws for Pre-training Agents and World Models7 Nov 2024 0 repositories listed
-
VideoGLaMM: A Large Multimodal Model for Pixel-Level Visual Grounding in Videos7 Nov 2024 0 repositories listed
-
VTechAGP: An Academic-to-General-Audience Text Paraphrase Dataset and Benchmark Models7 Nov 2024 0 repositories listed
-
Watermarking Language Models through Language Models7 Nov 2024 0 repositories listed
-
Deploying Multi-task Online Server with Large Language Model6 Nov 2024 0 repositories listed
-
Fine-Tuning Vision-Language Model for Automated Engineering Drawing Information Extraction6 Nov 2024 0 repositories listed
-
Large Generative Model-assisted Talking-face Semantic Communication System6 Nov 2024 0 repositories listed
-
NeurIPS 2023 Competition: Privacy Preserving Federated Learning Document VQA6 Nov 2024 0 repositories listed
-
The N-Grammys: Accelerating Autoregressive Inference with Learning-Free Batched Speculation6 Nov 2024 0 repositories listed
-
AI Metropolis: Scaling Large Language Model-based Multi-Agent Simulation with Out-of-order Execution5 Nov 2024 0 repositories listed
-
ChatGPT in Research and Education: Exploring Benefits and Threats5 Nov 2024 0 repositories listed
-
Controlling for Unobserved Confounding with Large Language Model Classification of Patient Smoking Status5 Nov 2024 0 repositories listed
-
HumanVLM: Foundation for Human-Scene Vision-Language Model5 Nov 2024 0 repositories listed
-
PersianRAG: A Retrieval-Augmented Generation System for Persian Language5 Nov 2024 0 repositories listed
-
Predictor-Corrector Enhanced Transformers with Exponential Moving Average Coefficient Learning5 Nov 2024 0 repositories listed
-
Spontaneous Emergence of Agent Individuality through Social Interactions in LLM-Based Communities5 Nov 2024 0 repositories listed
-
The Evolution of RWKV: Advancements in Efficient Language Modeling5 Nov 2024 0 repositories listed
-
Unified Pathological Speech Analysis with Prompt Tuning5 Nov 2024 0 repositories listed
-
[Vision Paper] PRObot: Enhancing Patient-Reported Outcome Measures for Diabetic Retinopathy using Chatbots and Generative AI5 Nov 2024 0 repositories listed
-
AVSS: Layer Importance Evaluation in Large Language Models via Activation Variance-Sparsity Analysis4 Nov 2024 0 repositories listed
-
ChatTracker: Enhancing Visual Tracking Performance via Chatting with Multimodal Large Language Model4 Nov 2024 0 repositories listed
-
Context Parallelism for Scalable Million-Token Inference4 Nov 2024 0 repositories listed
-
GraphVL: Graph-Enhanced Semantic Modeling via Vision-Language Models for Generalized Class Discovery4 Nov 2024 0 repositories listed
-
KptLLM: Unveiling the Power of Large Language Model for Keypoint Comprehension4 Nov 2024 0 repositories listed
-
Wave Network: An Ultra-Small Language Model4 Nov 2024 0 repositories listed
-
A Deep Dive Into Large Language Model Code Generation Mistakes: What and Why?3 Nov 2024 0 repositories listed
-
Enriching Tabular Data with Contextual LLM Embeddings: A Comprehensive Ablation Study for Ensemble Classifiers3 Nov 2024 0 repositories listed
-
High-performance automated abstract screening with large language model ensembles3 Nov 2024 0 repositories listed
-
Large Language Model Supply Chain: Open Problems From the Security Perspective3 Nov 2024 0 repositories listed
-
A Mechanistic Explanatory Strategy for XAI2 Nov 2024 0 repositories listed
-
Privacy Leakage Overshadowed by Views of AI: A Study on Human Oversight of Privacy in Language Model Agent2 Nov 2024 0 repositories listed
-
Can Large Language Model Predict Employee Attrition?2 Nov 2024 0 repositories listed
-
Can Multimodal Large Language Model Think Analogically?2 Nov 2024 0 repositories listed
-
Interacting Large Language Model Agents. Interpretable Models and Social Learning2 Nov 2024 0 repositories listed
-
PRIMO: Progressive Induction for Multi-hop Open Rule Generation2 Nov 2024 0 repositories listed
-
Swan and ArabicMTEB: Dialect-Aware, Arabic-Centric, Cross-Lingual, and Cross-Cultural Embedding Models and Benchmarks2 Nov 2024 0 repositories listed
-
Adding Error Bars to Evals: A Statistical Approach to Language Model Evaluations1 Nov 2024 0 repositories listed
-
Enhancing AAC Software for Dysarthric Speakers in e-Health Settings: An Evaluation Using TORGO1 Nov 2024 0 repositories listed
-
Enhancing the Traditional Chinese Medicine Capabilities of Large Language Model through Reinforcement Learning from AI Feedback1 Nov 2024 0 repositories listed
-
Improving Few-Shot Cross-Domain Named Entity Recognition by Instruction Tuning a Word-Embedding based Retrieval Augmented Large Language Model1 Nov 2024 0 repositories listed
-
LLM-KT: A Versatile Framework for Knowledge Transfer from Large Language Models to Collaborative Filtering1 Nov 2024 0 repositories listed
-
RadFlag: A Black-Box Hallucination Detection Method for Medical Vision Language Models1 Nov 2024 0 repositories listed
-
ReSpAct: Harmonizing Reasoning, Speaking, and Acting Towards Building Large Language Model-Based Conversational AI Agents1 Nov 2024 0 repositories listed
-
1 Nov 2024 0 repositories listed
-
Unified Generative and Discriminative Training for Multi-modal Large Language Models1 Nov 2024 0 repositories listed
-
ALISE: Accelerating Large Language Model Serving with Speculative Scheduling31 Oct 2024 0 repositories listed
-
Beyond Label Attention: Transparency in Language Models for Automated Medical Coding via Dictionary Learning31 Oct 2024 0 repositories listed
-
DEREC-SIMPRO: unlock Language Model benefits to advance Synthesis in Data Clean Room31 Oct 2024 0 repositories listed
-
From Context to Action: Analysis of the Impact of State Representation and Context on the Generalization of Multi-Turn Web Navigation Agents31 Oct 2024 0 repositories listed
-
31 Oct 2024 0 repositories listed Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
MESS+: Energy-Optimal Inferencing in Language Model Zoos with Service Level Guarantees31 Oct 2024 0 repositories listed
-
Morphological Typology in BPE Subword Productivity and Language Modeling31 Oct 2024 0 repositories listed
-
π₀: A Vision-Language-Action Flow Model for General Robot Control31 Oct 2024 0 repositories listed
-
Representative Social Choice: From Learning Theory to AI Alignment31 Oct 2024 0 repositories listed
-
Schema Augmentation for Zero-Shot Domain Adaptation in Dialogue State Tracking31 Oct 2024 0 repositories listed
-
Stereo-Talker: Audio-driven 3D Human Synthesis with Prior-Guided Mixture-of-Experts31 Oct 2024 0 repositories listed
-
The NPU-HWC System for the ISCSLP 2024 Inspirational and Convincing Audio Generation Challenge31 Oct 2024 0 repositories listed
-
Thought Space Explorer: Navigating and Expanding Thought Space for Large Language Model Reasoning31 Oct 2024 0 repositories listed
-
Towards Reliable Alignment: Uncertainty-aware RLHF31 Oct 2024 0 repositories listed
-
Web-Scale Visual Entity Recognition: An LLM-Driven Data Approach31 Oct 2024 0 repositories listed
-
Weight decay induces low-rank attention layers31 Oct 2024 0 repositories listed
-
A Monte Carlo Framework for Calibrated Uncertainty Estimation in Sequence Prediction30 Oct 2024 0 repositories listed
-
A Theoretical Perspective for Speculative Decoding Algorithm30 Oct 2024 0 repositories listed
-
All or None: Identifiable Linear Properties of Next-token Predictors in Language Modeling30 Oct 2024 0 repositories listed
-
Constructing Multimodal Datasets from Scratch for Rapid Development of a Japanese Visual Language Model30 Oct 2024 0 repositories listed
-
Dynamic Information Sub-Selection for Decision Support30 Oct 2024 0 repositories listed
-
Explainable Behavior Cloning: Teaching Large Language Model Agents through Learning by Demonstration30 Oct 2024 0 repositories listed
-
IP-MOT: Instance Prompt Learning for Cross-Domain Multi-Object Tracking30 Oct 2024 0 repositories listed
-
Learning and Transferring Sparse Contextual Bigrams with Linear Transformers30 Oct 2024 0 repositories listed
-
Prove Your Point!: Bringing Proof-Enhancement Principles to Argumentative Essay Generation30 Oct 2024 0 repositories listed
-
30 Oct 2024 0 repositories listed
-
Robotic State Recognition with Image-to-Text Retrieval Task of Pre-Trained Vision-Language Model and Black-Box Optimization30 Oct 2024 0 repositories listed
-
Smaller Large Language Models Can Do Moral Self-Correction30 Oct 2024 0 repositories listed
-
Teaching a Language Model to Distinguish Between Similar Details using a Small Adversarial Training Set30 Oct 2024 0 repositories listed
-
30 Oct 2024 0 repositories listed Syntology 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
VisualPredicator: Learning Abstract World Models with Neuro-Symbolic Predicates for Robot Planning30 Oct 2024 0 repositories listed
-
A Hierarchical Language Model For Interpretable Graph Reasoning29 Oct 2024 0 repositories listed
-
29 Oct 2024 0 repositories listed Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Anticipating Future with Large Language Model for Simultaneous Machine Translation29 Oct 2024 0 repositories listed
-
Auto-Intent: Automated Intent Discovery and Self-Exploration for Large Language Model Web Agents29 Oct 2024 0 repositories listed
-
CurateGPT: A flexible language-model assisted biocuration tool29 Oct 2024 0 repositories listed
-
Democratizing Reward Design for Personal and Representative Value-Alignment29 Oct 2024 0 repositories listed
-
Discrete Modeling via Boundary Conditional Diffusion Processes29 Oct 2024 0 repositories listed
-
FactBench: A Dynamic Benchmark for In-the-Wild Language Model Factuality Evaluation29 Oct 2024 0 repositories listed
-
From melodic note sequences to pitches using word2vec29 Oct 2024 0 repositories listed
-
Learning and Unlearning of Fabricated Knowledge in Language Models29 Oct 2024 0 repositories listed
-
MARCO: Multi-Agent Real-time Chat Orchestration29 Oct 2024 0 repositories listed
-
MotionGPT-2: A General-Purpose Motion-Language Model for Motion Generation and Understanding29 Oct 2024 0 repositories listed
-
Reliable Semantic Understanding for Real World Zero-shot Object Goal Navigation29 Oct 2024 0 repositories listed
-
VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration29 Oct 2024 0 repositories listed
-
An Actor-Critic Approach to Boosting Text-to-SQL Large Language Model28 Oct 2024 0 repositories listed
-
BongLLaMA: LLaMA for Bangla Language28 Oct 2024 0 repositories listed
-
Can Machines Think Like Humans? A Behavioral Evaluation of LLM-Agents in Dictator Games28 Oct 2024 0 repositories listed
Syntology lines on 3 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.