Browse State-of-the-Art › Large Language Model › Papers, page 25
Large Language Model
Papers archive 2025-07-28
archive papers tagged: 6,097 · with a code link: 2,250 · where Syntology ran a sample: 801 (683 with a run with no instrument failure, 118 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (801 of 6,097 tagged: 683 with a run with no instrument failure, 118 where every run was a failure of Syntology's instrument)
Page 25 of 61: papers 2,401 to 2,500 of 6,097, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
DeepVideo-R1: Video Reinforcement Fine-Tuning via Difficulty-aware Regressive GRPO9 Jun 2025 0 repositories listed
-
Event-Priori-Based Vision-Language Model for Efficient Visual Understanding9 Jun 2025 0 repositories listed
-
JavelinGuard: Low-Cost Transformer Architectures for LLM Security9 Jun 2025 0 repositories listed
-
Language-Grounded Hierarchical Planning and Execution with Multi-Robot 3D Scene Graphs9 Jun 2025 0 repositories listed
-
LLM Unlearning Should Be Form-Independent9 Jun 2025 0 repositories listed
-
QA-LIGN: Aligning LLMs through Constitutionally Decomposed QA9 Jun 2025 0 repositories listed
-
SpatialLM: Training Large Language Models for Structured Indoor Modeling9 Jun 2025 0 repositories listed
-
Statistical Hypothesis Testing for Auditing Robustness in Language Models9 Jun 2025 0 repositories listed
-
Speech Recognition on TV Series with Video-guided Post-Correction8 Jun 2025 0 repositories listed
-
An Agentic Framework for Autonomous Metamaterial Modeling and Inverse Design7 Jun 2025 0 repositories listed
-
Contextual Experience Replay for Self-Improvement of Language Agents7 Jun 2025 0 repositories listed
-
7 Jun 2025 0 repositories listed Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Voice Impression Control in Zero-Shot TTS6 Jun 2025 0 repositories listed
-
AgentSwift: Efficient LLM Agent Design via Value-guided Hierarchical Search6 Jun 2025 0 repositories listed
-
Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage6 Jun 2025 0 repositories listed
-
Hierarchical Debate-Based Large Language Model (LLM) for Complex Task Planning of 6G Network Management6 Jun 2025 0 repositories listed
-
PersonaAgent: When Large Language Model Agents Meet Personalization at Test Time6 Jun 2025 0 repositories listed
-
Prompting Wireless Networks: Reinforced In-Context Learning for Power Control6 Jun 2025 0 repositories listed
-
ScriptDoctor: Automatic Generation of PuzzleScript Games via Large Language Models and Tree Search6 Jun 2025 0 repositories listed
-
SPARQ: Synthetic Problem Generation for Reasoning via Quality-Diversity Algorithms6 Jun 2025 0 repositories listed
-
Training-Free Query Optimization via LLM-Based Plan Similarity6 Jun 2025 0 repositories listed
-
Subjective Perspectives within Learned Representations Predict High-Impact Innovation5 Jun 2025 0 repositories listed
-
Clustering and Median Aggregation Improve Differentially Private Inference5 Jun 2025 0 repositories listed
-
Customizing Speech Recognition Model with Large Language Model Feedback5 Jun 2025 0 repositories listed
-
DIMCIM: A Quantitative Evaluation Framework for Default-mode Diversity and Generalization in Text-to-Image Generative Models5 Jun 2025 0 repositories listed
-
E-bike agents: Large Language Model-Driven E-Bike Accident Analysis and Severity Prediction5 Jun 2025 0 repositories listed
-
Hierarchical Language Models for Semantic Navigation and Manipulation in an Aerial-Ground Robotic System5 Jun 2025 0 repositories listed
-
Interpretable Multimodal Framework for Human-Centered Street Assessment: Integrating Visual-Language Models for Perceptual Urban Diagnostics5 Jun 2025 0 repositories listed
-
LESS: Large Language Model Enhanced Semi-Supervised Learning for Speech Foundational Models5 Jun 2025 0 repositories listed
-
Parking, Perception, and Retail: Street-Level Determinants of Community Vitality in Harbin5 Jun 2025 0 repositories listed
-
Sparse Autoencoders, Again?5 Jun 2025 0 repositories listed
-
The NTNU System at the S&I Challenge 2025 SLA Open Track5 Jun 2025 0 repositories listed
-
Towards LLM-Centric Multimodal Fusion: A Survey on Integration Strategies and Techniques5 Jun 2025 0 repositories listed
-
A Novel Data Augmentation Approach for Automatic Speaking Assessment on Opinion Expressions4 Jun 2025 0 repositories listed
-
AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents4 Jun 2025 0 repositories listed
-
Automated Skill Discovery for Language Agents through Exploration and Iterative Feedback4 Jun 2025 0 repositories listed
-
CogniPair: From LLM Chatbots to Conscious AI Agents -- GNWT-Based Multi-Agent Digital Twins for Social Pairing -- Dating & Hiring Applications4 Jun 2025 0 repositories listed
-
"Don't Do That!": Guiding Embodied Systems through Large Language Model-based Constraint Generation4 Jun 2025 0 repositories listed
-
EuroLLM-9B: Technical Report4 Jun 2025 0 repositories listed
-
Evaluating Apple Intelligence's Writing Tools for Privacy Against Large Language Model-Based Inference Attacks: Insights from Early Datasets4 Jun 2025 0 repositories listed
-
Evaluating Large Language Model Capabilities in Assessing Spatial Econometrics Research4 Jun 2025 0 repositories listed
-
GEM: Empowering LLM for both Embedding Generation and Language Understanding4 Jun 2025 0 repositories listed
-
MedAgentGym: Training LLM Agents for Code-Based Medical Reasoning at Scale4 Jun 2025 0 repositories listed
-
The Cost of Dynamic Reasoning: Demystifying AI Agents and Test-Time Scaling from an AI Infrastructure Perspective4 Jun 2025 0 repositories listed
-
Understanding and Meeting Practitioner Needs When Measuring Representational Harms Caused by LLM-Based Systems4 Jun 2025 0 repositories listed
-
TaxAgent: How Large Language Model Designs Fiscal Policy3 Jun 2025 0 repositories listed
-
MASTER: Enhancing Large Language Model via Multi-Agent Simulated Teaching3 Jun 2025 0 repositories listed
-
TalkingMachines: Real-Time Audio-Driven FaceTime-Style Video via Autoregressive Diffusion Models3 Jun 2025 0 repositories listed
-
TestAgent: An Adaptive and Intelligent Expert for Human Assessment3 Jun 2025 0 repositories listed
-
COALESCE: Economic and Security Dynamics of Skill-Based Task Outsourcing Among Team of Autonomous LLM Agents2 Jun 2025 0 repositories listed
-
From Street Views to Urban Science: Discovering Road Safety Factors with Multimodal Large Language Models2 Jun 2025 0 repositories listed
-
Hybrid AI for Responsive Multi-Turn Online Conversations with Novel Dynamic Routing and Feedback Adaptation2 Jun 2025 0 repositories listed
-
Image Generation from Contextually-Contradictory Prompts2 Jun 2025 0 repositories listed
-
KDRL: Post-Training Reasoning LLMs via Unified Knowledge Distillation and Reinforcement Learning2 Jun 2025 0 repositories listed
-
LAM SIMULATOR: Advancing Data Generation for Large Action Model Training via Online Exploration and Trajectory Feedback2 Jun 2025 0 repositories listed
-
MLorc: Momentum Low-rank Compression for Large Language Model Adaptation2 Jun 2025 0 repositories listed
-
PGPO: Enhancing Agent Reasoning via Pseudocode-style Planning Guided Preference Optimization2 Jun 2025 0 repositories listed
-
PointT2I: LLM-based text-to-image generation via keypoints2 Jun 2025 0 repositories listed
-
WebChoreArena: Evaluating Web Browsing Agents on Realistic Tedious Web Tasks2 Jun 2025 0 repositories listed
-
Why Gradients Rapidly Increase Near the End of Training2 Jun 2025 0 repositories listed
-
A Large Language Model-Supported Threat Modeling Framework for Transportation Cyber-Physical Systems1 Jun 2025 0 repositories listed
-
Bridging Subjective and Objective QoE: Operator-Level Aggregation Using LLM-Based Comment Analysis and Network MOS Comparison1 Jun 2025 0 repositories listed
-
EEG2TEXT-CN: An Exploratory Study of Open-Vocabulary Chinese Text-EEG Alignment via Large Language Model and Contrastive Learning on ChineseEEG1 Jun 2025 0 repositories listed
-
HADA: Human-AI Agent Decision Alignment Architecture1 Jun 2025 0 repositories listed
-
Mamba Drafters for Speculative Decoding1 Jun 2025 0 repositories listed
-
OG-VLA: 3D-Aware Vision Language Action Model via Orthographic Image Generation1 Jun 2025 0 repositories listed
-
Organizational Adaptation to Generative AI in Cybersecurity: A Systematic Review31 May 2025 0 repositories listed
-
A Red Teaming Roadmap Towards System-Level Safety30 May 2025 0 repositories listed
-
A Reward-driven Automated Webshell Malicious-code Generator for Red-teaming30 May 2025 0 repositories listed
-
Artificial Empathy: AI based Mental Health30 May 2025 0 repositories listed
-
Beyond Exponential Decay: Rethinking Error Accumulation in Large Language Models30 May 2025 0 repositories listed
-
CREFT: Sequential Multi-Agent LLM for Character Relation Extraction30 May 2025 0 repositories listed
-
From Macro to Micro: Probing Dataset Diversity in Language Model Fine-Tuning30 May 2025 0 repositories listed
-
Grid-LOGAT: Grid Based Local and Global Area Transcription for Video Question Answering30 May 2025 0 repositories listed
-
HardTests: Synthesizing High-Quality Test Cases for LLM Coding30 May 2025 0 repositories listed
-
Hierarchical Level-Wise News Article Clustering via Multilingual Matryoshka Embeddings30 May 2025 0 repositories listed
-
Intuitionistic Fuzzy Sets for Large Language Model Data Annotation: A Novel Approach to Side-by-Side Preference Labeling30 May 2025 0 repositories listed
-
MythTriage: Scalable Detection of Opioid Use Disorder Myths on a Video-Sharing Platform30 May 2025 0 repositories listed
-
RoboMoRe: LLM-based Robot Co-design via Joint Optimization of Morphology and Reward30 May 2025 0 repositories listed
-
S4-Driver: Scalable Self-Supervised Driving Multimodal Large Language Modelwith Spatio-Temporal Visual Representation30 May 2025 0 repositories listed
-
SentinelAgent: Graph-based Anomaly Detection in Multi-Agent Systems30 May 2025 0 repositories listed
-
When GPT Spills the Tea: Comprehensive Assessment of Knowledge File Leakage in GPTs30 May 2025 0 repositories listed
-
Active Layer-Contrastive Decoding Reduces Hallucination in Large Language Model Generation29 May 2025 0 repositories listed
-
Cross-Task Experiential Learning on LLM-based Multi-Agent Collaboration29 May 2025 0 repositories listed
-
Dataset Cartography for Large Language Model Alignment: Mapping and Diagnosing Preference Data29 May 2025 0 repositories listed
-
Deep Retrieval at CheckThat! 2025: Identifying Scientific Papers from Implicit Social Media Mentions via Hybrid Retrieval and Re-Ranking29 May 2025 0 repositories listed
-
Diversity-Aware Policy Optimization for Large Language Model Reasoning29 May 2025 0 repositories listed
-
Evaluating the Efficacy of LLM-Based Reasoning for Multiobjective HPC Job Scheduling29 May 2025 0 repositories listed
-
Large Language Model-Based Agents for Automated Research Reproducibility: An Exploratory Study in Alzheimer's Disease29 May 2025 0 repositories listed
-
Large Language Model Meets Constraint Propagation29 May 2025 0 repositories listed
-
LLM Agents Should Employ Security Principles29 May 2025 0 repositories listed
-
Revisiting Multi-Agent Debate as Test-Time Scaling: A Systematic Study of Conditional Effectiveness29 May 2025 0 repositories listed
-
SCORPIO: Serving the Right Requests at the Right Time for Heterogeneous SLOs in LLM Inference29 May 2025 0 repositories listed
-
TailorSQL: An NL2SQL System Tailored to Your Query Workload29 May 2025 0 repositories listed
-
3DLLM-Mem: Long-Term Spatial-Temporal Memory for Embodied 3D Large Language Model28 May 2025 0 repositories listed
-
A Large Language Model-Enabled Control Architecture for Dynamic Resource Capability Exploration in Multi-Agent Manufacturing Systems28 May 2025 0 repositories listed
-
Agent-UniRAG: A Trainable Open-Source LLM Agent Framework for Unified Retrieval-Augmented Generation Systems28 May 2025 0 repositories listed
-
BugWhisperer: Fine-Tuning LLMs for SoC Hardware Vulnerability Detection28 May 2025 0 repositories listed
-
Conversational Alignment with Artificial Intelligence in Context28 May 2025 0 repositories listed
-
Design and testing of an agent chatbot supporting decision making with public transport data28 May 2025 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.