Browse State-of-the-Art › Language Modeling › Papers, page 88
Language Modeling
Papers archive 2025-07-28
archive papers tagged: 14,182 · with a code link: 5,620 · where Syntology ran a sample: 1,894 (1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,894 of 14,182 tagged: 1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument)
Page 88 of 142: papers 8,701 to 8,800 of 14,182, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
DYNA: Disease-Specific Language Model for Variant Pathogenicity31 May 2024 0 repositories listed
-
Exploratory Preference Optimization: Harnessing Implicit Q*-Approximation for Sample-Efficient RLHF31 May 2024 0 repositories listed
-
FineRadScore: A Radiology Report Line-by-Line Evaluation Technique Generating Corrections with Severity Scores31 May 2024 0 repositories listed
-
Kaleido Diffusion: Improving Conditional Diffusion Models with Autoregressive Latent Modeling31 May 2024 0 repositories listed
-
LOLAMEME: Logic, Language, Memory, Mechanistic Framework31 May 2024 0 repositories listed
-
Masked Language Modeling Becomes Conditional Density Estimation for Tabular Data Synthesis31 May 2024 0 repositories listed
-
RAG Does Not Work for Enterprises31 May 2024 0 repositories listed
-
StrucTexTv3: An Efficient Vision-Language Model for Text-rich Image Perception, Comprehension, and Beyond31 May 2024 0 repositories listed
-
You Only Scan Once: Efficient Multi-dimension Sequential Modeling with LightNet31 May 2024 0 repositories listed
-
30 May 2024 0 repositories listed
-
Efficient Indirect LLM Jailbreak via Multimodal-LLM Jailbreak30 May 2024 0 repositories listed
-
Knowledge Graph Tuning: Real-time Large Language Model Personalization based on Human Feedback30 May 2024 0 repositories listed
-
Knowledge-grounded Adaptation Strategy for Vision-language Models: Building Unique Case-set for Screening Mammograms for Residents Training30 May 2024 0 repositories listed
-
Large Language Model Watermark Stealing With Mixed Integer Programming30 May 2024 0 repositories listed
-
SeamlessExpressiveLM: Speech Language Model for Expressive Speech-to-Speech Translation with Chain-of-Thought30 May 2024 0 repositories listed
-
Who Writes the Review, Human or AI?30 May 2024 0 repositories listed
-
A Full-duplex Speech Dialogue Scheme Based On Large Language Models29 May 2024 0 repositories listed
-
Contextual Position Encoding: Learning to Count What's Important29 May 2024 0 repositories listed
-
Kotlin ML Pack: Technical Report29 May 2024 0 repositories listed
-
Learning from Litigation: Graphs and LLMs for Retrieval and Reasoning in eDiscovery29 May 2024 0 repositories listed
-
LLaMA-Reg: Using LLaMA 2 for Unsupervised Medical Image Registration29 May 2024 0 repositories listed
-
MindSemantix: Deciphering Brain Visual Experiences with a Brain-Language Model29 May 2024 0 repositories listed
-
Multi-Modal Generative Embedding Model29 May 2024 0 repositories listed
-
Nearest Neighbor Speculative Decoding for LLM Generation and Attribution29 May 2024 0 repositories listed
-
X-VILA: Cross-Modality Alignment for Large Language Model29 May 2024 0 repositories listed
-
Automated Real-World Sustainability Data Generation from Images of Buildings28 May 2024 0 repositories listed
-
Black-Box Detection of Language Model Watermarks28 May 2024 0 repositories listed
-
A Context-Aware Approach for Enhancing Data Imputation with Pre-trained Language Models28 May 2024 0 repositories listed
-
Don't Forget to Connect! Improving RAG with Graph-based Reranking28 May 2024 0 repositories listed
-
Facilitating Holistic Evaluations with LLMs: Insights from Scenario-Based Experiments28 May 2024 0 repositories listed
-
Unified Preference Optimization: Language Model Alignment Beyond the Preference Frontier28 May 2024 0 repositories listed
-
IAPT: Instruction-Aware Prompt Tuning for Large Language Models28 May 2024 0 repositories listed
-
Semantic are Beacons: A Semantic Perspective for Unveiling Parameter-Efficient Fine-Tuning in Knowledge Learning28 May 2024 0 repositories listed
-
Towards a theory of how the structure of language is acquired by deep neural networks28 May 2024 0 repositories listed
-
XL3M: A Training-free Framework for LLM Length Extension Based on Segment-wise Inference28 May 2024 0 repositories listed
-
A Large Language Model-based multi-agent manufacturing system for intelligent shopfloor27 May 2024 0 repositories listed
-
An Introduction to Vision-Language Modeling27 May 2024 0 repositories listed
-
Benchmarking General-Purpose In-Context Learning27 May 2024 0 repositories listed
-
LARM: Large Auto-Regressive Model for Long-Horizon Embodied Intelligence27 May 2024 0 repositories listed
-
Salutary Labeling with Zero Human Annotation27 May 2024 0 repositories listed
-
Self-Corrected Multimodal Large Language Model for End-to-End Robot Manipulation27 May 2024 0 repositories listed
-
SelfCP: Compressing Over-Limit Prompt via the Frozen Large Language Model Itself27 May 2024 0 repositories listed
-
SMR: State Memory Replay for Long Sequence Modeling27 May 2024 0 repositories listed
-
The Economic Implications of Large Language Model Selection on Earnings and Return on Investment: A Decision Theoretic Model27 May 2024 0 repositories listed
-
Unlocking the Secrets of Linear Complexity Sequence Model from A Unified Perspective27 May 2024 0 repositories listed
-
Chain of Tools: Large Language Model is an Automatic Multi-tool Learner26 May 2024 0 repositories listed
-
Code Repair with LLMs gives an Exploration-Exploitation Tradeoff26 May 2024 0 repositories listed
-
M-RAG: Reinforcing Large Language Model Performance through Retrieval-Augmented Generation with Multiple Partitions26 May 2024 0 repositories listed
-
Synthesizing Programmatic Reinforcement Learning Policies with Large Language Model Guided Search26 May 2024 0 repositories listed
-
Towards Multi-Task Multi-Modal Models: A Video Generative Perspective26 May 2024 0 repositories listed
-
A transfer learning framework for weak-to-strong generalization25 May 2024 0 repositories listed
-
How Well Do Deep Learning Models Capture Human Concepts? The Case of the Typicality Effect25 May 2024 0 repositories listed
-
Semantic Importance-Aware Communications with Semantic Correction Using Large Language Models25 May 2024 0 repositories listed
-
Theoretical Analysis of Weak-to-Strong Generalization25 May 2024 0 repositories listed
-
Decoding at the Speed of Thought: Harnessing Parallel Decoding of Lexical Units for LLMs24 May 2024 0 repositories listed
-
DnA-Eval: Enhancing Large Language Model Evaluation through Decomposition and Aggregation24 May 2024 0 repositories listed
-
Enhancing Augmentative and Alternative Communication with Card Prediction and Colourful Semantics24 May 2024 0 repositories listed
-
GECKO: Generative Language Model for English, Code and Korean24 May 2024 0 repositories listed
-
iREPO: implicit Reward Pairwise Difference based Empirical Preference Optimization24 May 2024 0 repositories listed
-
Inverse-RLignment: Large Language Model Alignment from Demonstrations through Inverse Reinforcement Learning24 May 2024 0 repositories listed
-
Large Language Model (LLM) for Standard Cell Layout Design Optimization24 May 2024 0 repositories listed
-
Large Language Model Pruning24 May 2024 0 repositories listed
-
Large Language Model Sentinel: LLM Agent for Adversarial Purification24 May 2024 0 repositories listed
-
Learning Beyond Pattern Matching? Assaying Mathematical Understanding in LLMs24 May 2024 0 repositories listed
-
Off-the-shelf ChatGPT is a Good Few-shot Human Motion Predictor24 May 2024 0 repositories listed
-
RAEE: A Robust Retrieval-Augmented Early Exiting Framework for Efficient Inference24 May 2024 0 repositories listed
-
Scaling Laws for Discriminative Classification in Large Language Models24 May 2024 0 repositories listed
-
24 May 2024 0 repositories listed Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Aya 23: Open Weight Releases to Further Multilingual Progress23 May 2024 0 repositories listed
-
BiMix: A Bivariate Data Mixing Law for Language Model Pretraining23 May 2024 0 repositories listed
-
Emotion Identification for French in Written Texts: Considering their Modes of Expression as a Step Towards Text Complexity Analysis23 May 2024 0 repositories listed
-
Exploring the use of a Large Language Model for data extraction in systematic reviews: a rapid feasibility study23 May 2024 0 repositories listed
-
23 May 2024 0 repositories listed Syntology 10 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 1 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Lessons from the Trenches on Reproducible Evaluation of Language Models23 May 2024 0 repositories listed
-
Visual Echoes: A Simple Unified Transformer for Audio-Visual Generation23 May 2024 0 repositories listed
-
Worldwide Federated Training of Language Models23 May 2024 0 repositories listed
-
22 May 2024 0 repositories listed Syntology 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
AI-Assisted Assessment of Coding Practices in Modern Code Review22 May 2024 0 repositories listed
-
Contextualized Automatic Speech Recognition with Dynamic Vocabulary22 May 2024 0 repositories listed
-
FiDeLiS: Faithful Reasoning in Large Language Model for Knowledge Graph Question Answering22 May 2024 0 repositories listed
-
HighwayLLM: Decision-Making and Navigation in Highway Driving with RL-Informed Language Model22 May 2024 0 repositories listed
-
KU-DMIS at EHRSQL 2024:Generating SQL query via question templatization in EHR22 May 2024 0 repositories listed
-
Slaves to the Law of Large Numbers: An Asymptotic Equipartition Property for Perplexity in Generative Language Models22 May 2024 0 repositories listed
-
Thermodynamic Natural Gradient Descent22 May 2024 0 repositories listed
-
Context-Enhanced Video Moment Retrieval with Large Language Models21 May 2024 0 repositories listed
-
Towards Retrieval-Augmented Architectures for Image Captioning21 May 2024 0 repositories listed
-
KG-RAG: Bridging the Gap Between Knowledge and Creativity20 May 2024 0 repositories listed
-
Knowledge-enhanced Prompt Tuning for Dialogue-based Relation Extraction with Trigger and Label Semantic20 May 2024 0 repositories listed
-
Quantifying In-Context Reasoning Effects and Memorization Effects in LLMs20 May 2024 0 repositories listed
-
Scientific Hypothesis Generation by a Large Language Model: Laboratory Validation in Breast Cancer Treatment20 May 2024 0 repositories listed
-
STYLE: Improving Domain Transferability of Asking Clarification Questions in Large Language Model Powered Conversational Agents20 May 2024 0 repositories listed
-
CPS-LLM: Large Language Model based Safe Usage Plan Generator for Human-in-the-Loop Human-in-the-Plant Cyber-Physical System19 May 2024 0 repositories listed
-
DocReLM: Mastering Document Retrieval with Language Model19 May 2024 0 repositories listed
-
EmbSum: Leveraging the Summarization Capabilities of Large Language Models for Content-Based Recommendations19 May 2024 0 repositories listed
-
VR-GPT: Visual Language Model for Intelligent Virtual Reality Applications19 May 2024 0 repositories listed
-
Automating PTSD Diagnostics in Clinical Interviews: Leveraging Large Language Models for Trauma Assessments18 May 2024 0 repositories listed
-
The CAP Principle for LLM Serving: A Survey of Long-Context Large Language Model Serving18 May 2024 0 repositories listed
-
ActiveLLM: Large Language Model-based Active Learning for Textual Few-Shot Scenarios17 May 2024 0 repositories listed
Syntology lines on 5 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.