Browse State-of-the-Art › Language Modeling › Papers, page 67
Language Modeling
Papers archive 2025-07-28
archive papers tagged: 14,182 · with a code link: 5,620 · where Syntology ran a sample: 1,894 (1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,894 of 14,182 tagged: 1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument)
Page 67 of 142: papers 6,601 to 6,700 of 14,182, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Do Sparse Autoencoders Generalize? A Case Study of Answerability27 Feb 2025 0 repositories listed
-
From Retrieval to Generation: Comparing Different Approaches27 Feb 2025 0 repositories listed
-
GRACE: A Granular Benchmark for Evaluating Model Calibration against Human Calibration27 Feb 2025 0 repositories listed
-
KEDRec-LM: A Knowledge-distilled Explainable Drug Recommendation Large Language Model27 Feb 2025 0 repositories listed
-
Large Language Model Strategic Reasoning Evaluation through Behavioral Game Theory27 Feb 2025 0 repositories listed
-
M-LLM Based Video Frame Selection for Efficient Video Understanding27 Feb 2025 0 repositories listed
-
NANOGPT: A Query-Driven Large Language Model Retrieval-Augmented Generation System for Nanotechnology Research27 Feb 2025 0 repositories listed
-
SEKI: Self-Evolution and Knowledge Inspiration based Neural Architecture Search via Large Language Models27 Feb 2025 0 repositories listed
-
Sparse Auto-Encoder Interprets Linguistic Features in Large Language Models27 Feb 2025 0 repositories listed
-
Tokens for Learning, Tokens for Unlearning: Mitigating Membership Inference Attacks in Large Language Models via Dual-Purpose Training27 Feb 2025 0 repositories listed
-
UniCodec: Unified Audio Codec with Single Domain-Adaptive Codebook27 Feb 2025 0 repositories listed
-
ANPMI: Assessing the True Comprehension Capabilities of LLMs for Multiple Choice Questions26 Feb 2025 0 repositories listed
-
Conformal Linguistic Calibration: Trading-off between Factuality and Specificity26 Feb 2025 0 repositories listed
-
I Know What I Don't Know: Improving Model Cascades Through Confidence Tuning26 Feb 2025 0 repositories listed
-
Improving Representation Learning of Complex Critical Care Data with ICU-BERT26 Feb 2025 0 repositories listed
-
Kanana: Compute-efficient Bilingual Language Models26 Feb 2025 0 repositories listed
-
Nexus: An Omni-Perceptive And -Interactive Model for Language, Audio, And Vision26 Feb 2025 0 repositories listed
-
On the Importance of Text Preprocessing for Multimodal Representation Learning and Pathology Report Generation26 Feb 2025 0 repositories listed
-
Pathology Report Generation and Multimodal Representation Learning for Cutaneous Melanocytic Lesions26 Feb 2025 0 repositories listed
-
The Sharpness Disparity Principle in Transformers for Accelerating Language Model Pre-Training26 Feb 2025 0 repositories listed
-
Revealing Treatment Non-Adherence Bias in Clinical Machine Learning Using Large Language Models26 Feb 2025 0 repositories listed
-
from Benign import Toxic: Jailbreaking the Language Model via Adversarial Metaphors25 Feb 2025 0 repositories listed
-
A Combinatorial Identities Benchmark for Theorem Proving via Automated Theorem Generation25 Feb 2025 0 repositories listed
-
AfroXLMR-Comet: Multilingual Knowledge Distillation with Attention Matching for Low-Resource languages25 Feb 2025 0 repositories listed
-
AMPO: Active Multi-Preference Optimization25 Feb 2025 0 repositories listed
-
Broadening Discovery through Structural Models: Multimodal Combination of Local and Structural Properties for Predicting Chemical Features25 Feb 2025 0 repositories listed
-
Can LLMs Explain Themselves Counterfactually?25 Feb 2025 0 repositories listed
-
Faster, Cheaper, Better: Multi-Objective Hyperparameter Optimization for LLM and RAG Systems25 Feb 2025 0 repositories listed
-
Independent Mobility GPT (IDM-GPT): A Self-Supervised Multi-Agent Large Language Model Framework for Customized Traffic Mobility Analysis Using Machine Learning Models25 Feb 2025 0 repositories listed
-
Large Language Model Driven Agents for Simulating Echo Chamber Formation25 Feb 2025 0 repositories listed
-
LDGen: Enhancing Text-to-Image Synthesis via Large Language Model-Driven Language Representation25 Feb 2025 0 repositories listed
-
MindMem: Multimodal for Predicting Advertisement Memorability Using LLMs and Deep Learning25 Feb 2025 0 repositories listed
-
PyEvalAI: AI-assisted evaluation of Jupyter Notebooks for immediate personalized feedback25 Feb 2025 0 repositories listed
-
VALUE: Value-Aware Large Language Model for Query Rewriting via Weighted Trie in Sponsored Search25 Feb 2025 0 repositories listed
-
Your Language Model May Think Too Rigidly: Achieving Reasoning Consistency with Symmetry-Enhanced Training25 Feb 2025 0 repositories listed
-
An Enhanced Large Language Model For Cross Modal Query Understanding System Using DL-KeyBERT Based CAZSSCL-MPGPT24 Feb 2025 0 repositories listed
-
Balancing Speech Understanding and Generation Using Continual Pre-training for Codec-based Speech LLM24 Feb 2025 0 repositories listed
-
Forecasting Rare Language Model Behaviors24 Feb 2025 0 repositories listed
-
From Perceptions to Decisions: Wildfire Evacuation Decision Prediction with Behavioral Theory-informed LLMs24 Feb 2025 0 repositories listed
-
How Do Large Language Monkeys Get Their Power (Laws)?24 Feb 2025 0 repositories listed
-
IGDA: Interactive Graph Discovery through Large Language Model Agents24 Feb 2025 0 repositories listed
-
Improving Interactive Diagnostic Ability of a Large Language Model Agent Through Clinical Experience Learning24 Feb 2025 0 repositories listed
-
Knowledge Distillation with Training Wheels24 Feb 2025 0 repositories listed
-
Language Model Re-rankers are Steered by Lexical Similarities24 Feb 2025 0 repositories listed
-
Predicting Liquidity-Aware Bond Yields using Causal GANs and Deep Reinforcement Learning with LLM Evaluation24 Feb 2025 0 repositories listed
-
Real-time Monitoring of Economic Shocks using Company Websites24 Feb 2025 0 repositories listed
-
Reasoning with Latent Thoughts: On the Power of Looped Transformers24 Feb 2025 0 repositories listed
-
Sarang at DEFACTIFY 4.0: Detecting AI-Generated Text Using Noised Data and an Ensemble of DeBERTa Models24 Feb 2025 0 repositories listed
-
24 Feb 2025 0 repositories listed Syntology 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 6 unverified (of 12 harvested samples) · 12 pointer-only (licence)
-
Thus Spake Long-Context Large Language Model24 Feb 2025 0 repositories listed
-
Time series forecasting based on optimized LLM for fault prediction in distribution power grid insulators24 Feb 2025 0 repositories listed
-
Unbiased and Sign Compression in Distributed Learning: Comparing Noise Resilience via SDEs24 Feb 2025 0 repositories listed
-
Understanding the Uncertainty of LLM Explanations: A Perspective Based on Reasoning Topology24 Feb 2025 0 repositories listed
-
Zero-shot Load Forecasting for Integrated Energy Systems: A Large Language Model-based Framework with Multi-task Learning24 Feb 2025 0 repositories listed
-
Retrieval-Augmented Visual Question Answering via Built-in Autoregressive Search Engines23 Feb 2025 0 repositories listed
-
Sequence-level Large Language Model Training with Contrastive Preference Optimization23 Feb 2025 0 repositories listed
-
A Framework for Evaluating Vision-Language Model Safety: Building Trust in AI for Public Sector Applications22 Feb 2025 0 repositories listed
-
Dynamic Parallel Tree Search for Efficient LLM Reasoning22 Feb 2025 0 repositories listed
-
Echo: A Large Language Model with Temporal Episodic Memory22 Feb 2025 0 repositories listed
-
Exploring Sentiment Manipulation by LLM-Enabled Intelligent Trading Agents22 Feb 2025 0 repositories listed
-
Human Preferences in Large Language Model Latent Space: A Technical Analysis on the Reliability of Synthetic Data in Voting Outcome Prediction22 Feb 2025 0 repositories listed
-
Large Language Model for Lossless Image Compression with Visual Prompts22 Feb 2025 0 repositories listed
-
Prompt as Knowledge Bank: Boost Vision-language model via Structural Representation for zero-shot medical detection22 Feb 2025 0 repositories listed
-
QWENDY: Gene Regulatory Network Inference Enhanced by Large Language Model and Transformer22 Feb 2025 0 repositories listed
-
Recurrent Knowledge Identification and Fusion for Language Model Continual Learning22 Feb 2025 0 repositories listed
-
Understanding Zero-shot Rare Word Recognition Improvements Through LLM Integration22 Feb 2025 0 repositories listed
-
ZiGong 1.0: A Large Language Model for Financial Credit22 Feb 2025 0 repositories listed
-
Bridging vision language model (VLM) evaluation gaps with a framework for scalable and cost-effective benchmark generation21 Feb 2025 0 repositories listed
-
Chitrarth: Bridging Vision and Language for a Billion People21 Feb 2025 0 repositories listed
-
Coherency Improved Explainable Recommendation via Large Language Model21 Feb 2025 0 repositories listed
-
Forecasting Frontier Language Model Agent Capabilities21 Feb 2025 0 repositories listed
-
Identifying Features that Shape Perceived Consciousness in Large Language Model-based AI: A Quantitative Study of Human Responses21 Feb 2025 0 repositories listed
-
LEDD: Large Language Model-Empowered Data Discovery in Data Lakes21 Feb 2025 0 repositories listed
-
MOVE: A Mixture-of-Vision-Encoders Approach for Domain-Focused Vision-Language Processing21 Feb 2025 0 repositories listed
-
Optimizing Pre-Training Data Mixtures with Mixtures of Data Expert Models21 Feb 2025 0 repositories listed
-
PAPI: Exploiting Dynamic Parallelism in Large Language Model Decoding with a Processing-In-Memory-Enabled Computing System21 Feb 2025 0 repositories listed
-
R³Mem: Bridging Memory Retention and Retrieval via Reversible Compression21 Feb 2025 0 repositories listed
-
Exploring Advanced Techniques for Visual Question Answering: A Comprehensive Comparison20 Feb 2025 0 repositories listed
-
FR-Spec: Accelerating Large-Vocabulary Language Models via Frequency-Ranked Speculative Sampling20 Feb 2025 0 repositories listed
-
HPS: Hard Preference Sampling for Human Preference Alignment20 Feb 2025 0 repositories listed
-
Optimizing Singular Spectrum for Large Language Model Compression20 Feb 2025 0 repositories listed
-
20 Feb 2025 0 repositories listed Syntology 4 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Show Me Your Code! Kill Code Poisoning: A Lightweight Method Based on Code Naturalness20 Feb 2025 0 repositories listed
-
SR-LLM: Rethinking the Structured Representation in Large Language Model20 Feb 2025 0 repositories listed
-
A Chain-of-Thought Subspace Meta-Learning for Few-shot Image Captioning with Large Vision and Language Models19 Feb 2025 0 repositories listed
-
Autellix: An Efficient Serving Engine for LLM Agents as General Programs19 Feb 2025 0 repositories listed
-
Complex Ontology Matching with Large Language Model Embeddings19 Feb 2025 0 repositories listed
-
Diversity-driven Data Selection for Language Model Tuning through Sparse Autoencoder19 Feb 2025 0 repositories listed
-
Event Segmentation Applications in Large Language Model Enabled Automated Recall Assessments19 Feb 2025 0 repositories listed
-
Flow-based generative models as iterative algorithms in probability space19 Feb 2025 0 repositories listed
-
LLM should think and action as a human19 Feb 2025 0 repositories listed
-
Megrez-Omni Technical Report19 Feb 2025 0 repositories listed
-
Mol-LLaMA: Towards General Understanding of Molecules in Large Molecular Language Model19 Feb 2025 0 repositories listed
-
Reducing Hallucinations in Language Model-based SPARQL Query Generation Using Post-Generation Memory Retrieval19 Feb 2025 0 repositories listed
-
REFIND: Retrieval-Augmented Factuality Hallucination Detection in Large Language Models19 Feb 2025 0 repositories listed
-
Reflection of Episodes: Learning to Play Game from Expert and Self Experiences19 Feb 2025 0 repositories listed
-
UniKnow: A Unified Framework for Reliable Language Model Behavior across Parametric and External Knowledge19 Feb 2025 0 repositories listed
-
Remote Sensing Semantic Segmentation Quality Assessment based on Vision Language Model19 Feb 2025 0 repositories listed
-
Retrieving Versus Understanding Extractive Evidence in Few-Shot Learning19 Feb 2025 0 repositories listed
-
RGAR: Recurrence Generation-augmented Retrieval for Factual-aware Medical Question Answering19 Feb 2025 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.