Browse State-of-the-Art › Language Modeling › Papers, page 63
Language Modeling
Papers archive 2025-07-28
archive papers tagged: 14,182 · with a code link: 5,620 · where Syntology ran a sample: 1,894 (1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,894 of 14,182 tagged: 1,580 with a run with no instrument failure, 314 where every run was a failure of Syntology's instrument)
Page 63 of 142: papers 6,201 to 6,300 of 14,182, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding20 Apr 2025 0 repositories listed
-
PROMPTEVALS: A Dataset of Assertions and Guardrails for Custom Production Large Language Model Pipelines20 Apr 2025 0 repositories listed
-
ResNetVLLM-2: Addressing ResNetVLLM's Multi-Modal Hallucinations20 Apr 2025 0 repositories listed
-
ResNetVLLM -- Multi-modal Vision LLM for the Video Understanding Task20 Apr 2025 0 repositories listed
-
A Multimodal Recaptioning Framework to Account for Perceptual Diversity in Multilingual Vision-Language Modeling19 Apr 2025 0 repositories listed
-
Bottom-Up Synthesis of Knowledge-Grounded Task-Oriented Dialogues with Iteratively Self-Refined Prompts19 Apr 2025 0 repositories listed
-
Improving the Serving Performance of Multi-LoRA Large Language Models via Efficient LoRA and KV Cache Management19 Apr 2025 0 repositories listed
-
Large Language Model Enhanced Particle Swarm Optimization for Hyperparameter Tuning for Deep Learning Models19 Apr 2025 0 repositories listed
-
A Baseline for Self-state Identification and Classification in Mental Health Data: CLPsych 2025 Task18 Apr 2025 0 repositories listed
-
Chain-of-Thought Textual Reasoning for Few-shot Temporal Action Localization18 Apr 2025 0 repositories listed
-
Feature Alignment and Representation Transfer in Knowledge Distillation for Large Language Models18 Apr 2025 0 repositories listed
-
Large Language Bayes18 Apr 2025 0 repositories listed
-
PV-VLM: A Multimodal Vision-Language Approach Incorporating Sky Images for Intra-Hour Photovoltaic Power Forecasting18 Apr 2025 0 repositories listed
-
RAG Without the Lag: Interactive Debugging for Retrieval-Augmented Generation Pipelines18 Apr 2025 0 repositories listed
-
System of Agentic AI for the Discovery of Metal-Organic Frameworks18 Apr 2025 0 repositories listed
-
Are Retrials All You Need? Enhancing Large Language Model Reasoning Without Verbalized Feedback17 Apr 2025 0 repositories listed
-
ChatEXAONEPath: An Expert-level Multimodal Large Language Model for Histopathology Using Whole Slide Images17 Apr 2025 0 repositories listed
-
CLIMB: CLustering-based Iterative Data Mixture Bootstrapping for Language Model Pre-training17 Apr 2025 0 repositories listed
-
17 Apr 2025 0 repositories listed Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
It's All Connected: A Journey Through Test-Time Memorization, Attentional Bias, Retention, and Online Optimization17 Apr 2025 0 repositories listed
-
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training17 Apr 2025 0 repositories listed
-
Pandora: A Code-Driven Large Language Model Agent for Unified Reasoning Across Diverse Structured Knowledge17 Apr 2025 0 repositories listed
-
VLLFL: A Vision-Language Model Based Lightweight Federated Learning Framework for Smart Agriculture17 Apr 2025 0 repositories listed
-
BitNet b1.58 2B4T Technical Report16 Apr 2025 0 repositories listed
-
d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning16 Apr 2025 0 repositories listed
-
DVLTA-VQA: Decoupled Vision-Language Modeling with Text-Guided Adaptation for Blind Video Quality Assessment16 Apr 2025 0 repositories listed
-
Generative Recommendation with Continuous-Token Diffusion16 Apr 2025 0 repositories listed
-
Higher-Order Binding of Language Model Virtual Personas: a Study on Approximating Political Partisan Misperceptions16 Apr 2025 0 repositories listed
-
Interpreting the linear structure of vision-language model embedding spaces16 Apr 2025 0 repositories listed
-
Mixer Metaphors: audio interfaces for non-musical applications16 Apr 2025 0 repositories listed
-
Rethinking LLM-Based Recommendations: A Query Generation-Based, Training-Free Approach16 Apr 2025 0 repositories listed
-
Securing the Skies: A Comprehensive Survey on Anti-UAV Methods, Benchmarking, and Future Directions16 Apr 2025 0 repositories listed
-
Towards Conversational AI for Human-Machine Collaborative MLOps16 Apr 2025 0 repositories listed
-
A Large-Language Model Framework for Relative Timeline Extraction from PubMed Case Reports15 Apr 2025 0 repositories listed
-
Co-STAR: Collaborative Curriculum Self-Training with Adaptive Regularization for Source-Free Video Domain Adaptation15 Apr 2025 0 repositories listed
-
DeepMLF: Multimodal language model with learnable tokens for deep fusion in sentiment analysis15 Apr 2025 0 repositories listed
-
Efficient Distributed Retrieval-Augmented Generation for Enhancing Language Model Performance15 Apr 2025 0 repositories listed
-
Efficient Hybrid Language Model Compression through Group-Aware SSM Pruning15 Apr 2025 0 repositories listed
-
From Gaze to Insight: Bridging Human Visual Attention and Vision Language Model Explanation for Weakly-Supervised Medical Image Segmentation15 Apr 2025 0 repositories listed
-
GraphicBench: A Planning Benchmark for Graphic Design with Language Agents15 Apr 2025 0 repositories listed
-
Large Language Model-Informed Feature Discovery Improves Prediction and Interpretation of Credibility Perceptions of Visual Content15 Apr 2025 0 repositories listed
-
Looking beyond the next token15 Apr 2025 0 repositories listed
-
ProtFlow: Fast Protein Sequence Design via Flow Matching on Compressed Protein Language Model Embeddings15 Apr 2025 0 repositories listed
-
Recommending Clinical Trials for Online Patient Cases using Artificial Intelligence15 Apr 2025 0 repositories listed
-
ReZero: Enhancing LLM search ability by trying one-more-time15 Apr 2025 0 repositories listed
-
α-Flow: A Unified Framework for Continuous-State Discrete Flow Matching Models14 Apr 2025 0 repositories listed
-
A Survey of Large Language Model-Powered Spatial Intelligence Across Scales: Advances in Embodied Agents, Smart Cities, and Earth Science14 Apr 2025 0 repositories listed
-
Automated Testing of COBOL to Java Transformation14 Apr 2025 0 repositories listed
-
Benchmarking Practices in LLM-driven Offensive Security: Testbeds, Metrics, and Experiment Design14 Apr 2025 0 repositories listed
-
Forecasting from Clinical Textual Time Series: Adaptations of the Encoder and Decoder Language Model Families14 Apr 2025 0 repositories listed
-
GNN-ACLP: Graph Neural Networks based Analog Circuit Link Prediction14 Apr 2025 0 repositories listed
-
LangPert: Detecting and Handling Task-level Perturbations for Robust Object Rearrangement14 Apr 2025 0 repositories listed
-
Learning from Reference Answers: Versatile Language Model Alignment without Binary Human Preference Data14 Apr 2025 0 repositories listed
-
Mavors: Multi-granularity Video Representation for Multimodal Large Language Model14 Apr 2025 0 repositories listed
-
MorphTok: Morphologically Grounded Tokenization for Indian Languages14 Apr 2025 0 repositories listed
-
Pseudo-Autoregressive Neural Codec Language Models for Efficient Zero-Shot Text-to-Speech Synthesis14 Apr 2025 0 repositories listed
-
SlowFastVAD: Video Anomaly Detection via Integrating Simple Detector and RAG-Enhanced Vision-Language Model14 Apr 2025 0 repositories listed
-
AgentA/B: Automated and Scalable Web A/BTesting with Interactive LLM Agents13 Apr 2025 0 repositories listed
-
AgentDynEx: Nudging the Mechanics and Dynamics of Multi-Agent Simulations13 Apr 2025 0 repositories listed
-
Domain-Adaptive Continued Pre-Training of Small Language Models13 Apr 2025 0 repositories listed
-
Kongzi: A Historical Large Language Model with Fact Enhancement13 Apr 2025 0 repositories listed
-
Structure-Accurate Medical Image Translation via Dynamic Frequency Balance and Knowledge Guidance13 Apr 2025 0 repositories listed
-
UXAgent: A System for Simulating Usability Testing of Web Design with LLM Agents13 Apr 2025 0 repositories listed
-
AstroLLaVA: towards the unification of astronomical data and natural language11 Apr 2025 0 repositories listed
-
ELSA: A Style Aligned Dataset for Emotionally Intelligent Language Generation11 Apr 2025 0 repositories listed
-
EO-VLM: VLM-Guided Energy Overload Attacks on Vision Models11 Apr 2025 0 repositories listed
-
Large Language Model Empowered Recommendation Meets All-domain Continual Pre-Training11 Apr 2025 0 repositories listed
-
Spatial Audio Processing with Large Language Model on Wearable Devices11 Apr 2025 0 repositories listed
-
SpecEE: Accelerating Large Language Model Inference with Speculative Early Exiting11 Apr 2025 0 repositories listed
-
SWAN-GPT: An Efficient and Scalable Approach for Long-Context Language Modeling11 Apr 2025 0 repositories listed
-
TP-RAG: Benchmarking Retrieval-Augmented Large Language Model Agents for Spatiotemporal-Aware Travel Planning11 Apr 2025 0 repositories listed
-
An LLM-Driven Multi-Agent Debate System for Mendelian Diseases10 Apr 2025 0 repositories listed
-
Beyond LLMs: A Linguistic Approach to Causal Graph Generation from Narrative Texts10 Apr 2025 0 repositories listed
-
Cat, Rat, Meow: On the Alignment of Language Model and Human Term-Similarity Judgments10 Apr 2025 0 repositories listed
-
Data Metabolism: An Efficient Data Design Schema For Vision Language Model10 Apr 2025 0 repositories listed
-
DeepGreen: Effective LLM-Driven Green-washing Monitoring System Designed for Empirical Testing -- Evidence from China10 Apr 2025 0 repositories listed
-
Investigating Vision-Language Model for Point Cloud-based Vehicle Classification10 Apr 2025 0 repositories listed
-
JEPA4Rec: Learning Effective Language Representations for Sequential Recommendation via Joint Embedding Predictive Architecture10 Apr 2025 0 repositories listed
-
Synthetic Fluency: Hallucinations, Confabulations, and the Creation of Irish Words in LLM-Generated Translations10 Apr 2025 0 repositories listed
-
Token Level Routing Inference System for Edge Devices10 Apr 2025 0 repositories listed
-
A Multi-Phase Analysis of Blood Culture Stewardship: Machine Learning Prediction, Expert Recommendation Assessment, and LLM Automation9 Apr 2025 0 repositories listed
-
Language Modeling for the Future of Finance: A Quantitative Survey into Metrics, Tasks, and Data Opportunities9 Apr 2025 0 repositories listed
-
MovSAM: A Single-image Moving Object Segmentation Framework Based on Deep Thinking9 Apr 2025 0 repositories listed
-
OLMoTrace: Tracing Language Model Outputs Back to Trillions of Training Tokens9 Apr 2025 0 repositories listed
-
Q-Agent: Quality-Driven Chain-of-Thought Image Restoration Agent through Robust Multimodal Large Language Model9 Apr 2025 0 repositories listed
-
Societal Impacts Research Requires Benchmarks for Creative Composition Tasks9 Apr 2025 0 repositories listed
-
The Method for Storing Patterns in Neural Networks-Memorization and Recall of QR code Patterns-9 Apr 2025 0 repositories listed
-
InstructMPC: A Human-LLM-in-the-Loop Framework for Context-Aware Control8 Apr 2025 0 repositories listed
-
Simplifying Data Integration: SLM-Driven Systems for Unified Semantic Queries Across Heterogeneous Databases8 Apr 2025 0 repositories listed
-
A Taxonomy of Self-Handover7 Apr 2025 0 repositories listed
-
Evaluating Knowledge Graph Based Retrieval Augmented Generation Methods under Knowledge Incompleteness7 Apr 2025 0 repositories listed
-
Large Language Model (LLM) for Software Security: Code Analysis, Malware Analysis, Reverse Engineering7 Apr 2025 0 repositories listed
-
'Neural howlround' in large language models: a self-reinforcing bias phenomenon, and a dynamic attenuation solution7 Apr 2025 0 repositories listed
-
The Dream Within Huang Long Cave: AI-Driven Interactive Narrative for Family Storytelling and Emotional Reflection7 Apr 2025 0 repositories listed
-
Towards Visual Text Grounding of Multimodal Large Language Model7 Apr 2025 0 repositories listed
-
Unleashing the Power of LLMs in Dense Retrieval with Query Likelihood Modeling7 Apr 2025 0 repositories listed
-
DDPT: Diffusion-Driven Prompt Tuning for Large Language Model Code Generation6 Apr 2025 0 repositories listed
-
ZeroED: Hybrid Zero-shot Error Detection through Large Language Model Reasoning6 Apr 2025 0 repositories listed
-
Large Language Model-Based Knowledge Graph System Construction for Sustainable Development Goals: An AI-Based Speculative Design Perspective5 Apr 2025 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.