Browse State-of-the-Art › Hallucination › Papers, page 9
Hallucination
Papers archive 2025-07-28
archive papers tagged: 1,816 · with a code link: 752 · where Syntology ran a sample: 276 (240 with a run with no instrument failure, 36 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (276 of 1,816 tagged: 240 with a run with no instrument failure, 36 where every run was a failure of Syntology's instrument)
Page 9 of 19: papers 801 to 900 of 1,816, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Reinforcement Learning for Better Verbalized Confidence in Long-Form Generation29 May 2025 0 repositories listed
-
Evaluation Hallucination in Multi-Round Incomplete Information Lateral-Driven Reasoning Tasks28 May 2025 0 repositories listed
-
SkewRoute: Training-Free LLM Routing for Knowledge Graph Retrieval-Augmented Generation via Score Skewness of Retrieved Context28 May 2025 0 repositories listed
-
A Lightweight Multi-Expert Generative Language Model System for Engineering Information and Knowledge Extraction27 May 2025 0 repositories listed
-
Mitigating Hallucination in Large Vision-Language Models via Adaptive Attention Calibration27 May 2025 0 repositories listed
-
Enhancing Visual Reliance in Text Generation: A Bayesian Perspective on Mitigating Hallucination in Large Vision-Language Models26 May 2025 0 repositories listed
-
Grounding Language with Vision: A Conditional Mutual Information Calibrated Decoding Strategy for Reducing Hallucinations in LVLMs26 May 2025 0 repositories listed
-
Uncertainty-Aware Attention Heads: Efficient Unsupervised Uncertainty Quantification for LLMs26 May 2025 0 repositories listed
-
GUARDIAN: Safeguarding LLM Multi-Agent Collaborations with Temporal Graph Modeling25 May 2025 0 repositories listed
-
LLLMs: A Data-Driven Survey of Evolving Research on Limitations of Large Language Models25 May 2025 0 repositories listed
-
More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models23 May 2025 0 repositories listed
-
Teaching with Lies: Curriculum DPO on Synthetic Negatives for Hallucination Detection23 May 2025 0 repositories listed
-
Chain-of-Thought Poisoning Attacks against R1-based Retrieval-Augmented Generation Systems22 May 2025 0 repositories listed
-
LLM-Powered Agents for Navigating Venice's Historical Cadastre22 May 2025 0 repositories listed
-
Locate-then-Merge: Neuron-Level Parameter Fusion for Mitigating Catastrophic Forgetting in Multimodal LLMs22 May 2025 0 repositories listed
-
Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention Causal Decoding22 May 2025 0 repositories listed
-
Shadows in the Attention: Contextual Perturbation and Representation Drift in the Dynamics of Hallucination in LLMs22 May 2025 0 repositories listed
-
Steering LVLMs via Sparse Autoencoder for Hallucination Mitigation22 May 2025 0 repositories listed
-
UNCLE: Uncertainty Expressions in Long-Form Generation22 May 2025 0 repositories listed
-
Aug2Search: Enhancing Facebook Marketplace Search with LLM-Generated Synthetic Data Augmentation21 May 2025 0 repositories listed
-
Hallucinate at the Last in Long Response Generation: A Case Study on Long Document Summarization21 May 2025 0 repositories listed
-
HCRMP: A LLM-Hinted Contextual Reinforcement Learning Framework for Autonomous Driving21 May 2025 0 repositories listed
-
KaFT: Knowledge-aware Fine-tuning for Boosting LLMs' Domain-specific Question-Answering Performance21 May 2025 0 repositories listed
-
Multilingual Prompting for Improving LLM Generation Diversity21 May 2025 0 repositories listed
-
NEXT-EVAL: Next Evaluation of Traditional and LLM Web Data Record Extraction21 May 2025 0 repositories listed
-
OViP: Online Vision-Language Preference Learning21 May 2025 0 repositories listed
-
RePPL: Recalibrating Perplexity by Uncertainty in Semantic Propagation and Language Generation for Explainable QA Hallucination Detection21 May 2025 0 repositories listed
-
Aligning Attention Distribution to Information Flow for Hallucination Mitigation in Large Vision-Language Models20 May 2025 0 repositories listed
-
Foundations of Unknown-aware Machine Learning20 May 2025 0 repositories listed
-
JARVIS: A Multi-Agent Code Assistant for High-Quality EDA Script Generation20 May 2025 0 repositories listed
-
Legal Rule Induction: Towards Generalizable Principle Discovery from Analogous Judicial Precedents20 May 2025 0 repositories listed
-
Multimodal RAG-driven Anomaly Detection and Classification in Laser Powder Bed Fusion using Large Language Models20 May 2025 0 repositories listed
-
Plane Geometry Problem Solving with Multi-modal Reasoning: A Survey20 May 2025 0 repositories listed
-
Reinforcing Question Answering Agents with Minimalist Policy Gradient Optimization20 May 2025 0 repositories listed
-
The Hallucination Tax of Reinforcement Finetuning20 May 2025 0 repositories listed
-
Towards Omnidirectional Reasoning with 360-R1: A Dataset, Benchmark, and GRPO-based Method20 May 2025 0 repositories listed
-
Visual Instruction Bottleneck Tuning20 May 2025 0 repositories listed
-
Calm-Whisper: Reduce Whisper Hallucination On Non-Speech By Calming Crazy Heads Down19 May 2025 0 repositories listed
-
Detection and Mitigation of Hallucination in Large Reasoning Models: A Mechanistic Perspective19 May 2025 0 repositories listed
-
Granary: Speech Recognition and Translation Dataset in 25 European Languages19 May 2025 0 repositories listed
-
Mitigating Hallucination in VideoLLMs via Temporal-Aware Activation Engineering19 May 2025 0 repositories listed
-
Selective Code Generation for Functional Guarantees19 May 2025 0 repositories listed
-
Tianyi: A Traditional Chinese Medicine all-rounder language model and its Real-World Clinical Practice19 May 2025 0 repositories listed
-
Learning Auxiliary Tasks Improves Reference-Free Hallucination Detection in Open-Domain Long-Form Generation18 May 2025 0 repositories listed
-
Mitigating Hallucinations via Inter-Layer Consistency Aggregation in Large Vision-Language Models18 May 2025 0 repositories listed
-
The Tower of Babel Revisited: Multilingual Jailbreak Prompts on Closed-Source Large Language Models18 May 2025 0 repositories listed
-
Are Multimodal Large Language Models Ready for Omnidirectional Spatial Reasoning?17 May 2025 0 repositories listed
-
CCNU at SemEval-2025 Task 3: Leveraging Internal and External Knowledge of Large Language Models for Multilingual Hallucination Annotation17 May 2025 0 repositories listed
-
Diverging Towards Hallucination: Detection of Failures in Vision-Language Models via Multi-token Aggregation16 May 2025 0 repositories listed
-
Towards Robust Evaluation of STEM Education: Leveraging MLLMs in Project-Based Learning16 May 2025 0 repositories listed
-
AI Agents vs. Agentic AI: A Conceptual Taxonomy, Applications and Challenges15 May 2025 0 repositories listed
-
A Multimodal Multi-Agent Framework for Radiology Report Generation14 May 2025 0 repositories listed
-
Beyond the Black Box: Interpretability of LLMs in Finance14 May 2025 0 repositories listed
-
Ornithologist: Towards Trustworthy "Reasoning" about Central Bank Communications14 May 2025 0 repositories listed
-
The Impact of Large Language Models on Task Automation in Manufacturing Services14 May 2025 0 repositories listed
-
Adaptive Schema-aware Event Extraction with Retrieval-Augmented Generation13 May 2025 0 repositories listed
-
Improving the Reliability of LLMs: Combining CoT, RAG, Self-Consistency, and Self-Verification13 May 2025 0 repositories listed
-
Critique Before Thinking: Mitigating Hallucination through Rationale-Augmented Instruction Tuning12 May 2025 0 repositories listed
-
On the Cost and Benefits of Training Context with Utterance or Full Conversation Training: A Comparative Stud12 May 2025 0 repositories listed
-
SEReDeEP: Hallucination Detection in Retrieval-Augmented Models via Semantic Entropy and Context-Parameter Fusion12 May 2025 0 repositories listed
-
TrumorGPT: Graph-Based Retrieval-Augmented Large Language Model for Fact-Checking11 May 2025 0 repositories listed
-
Osiris: A Lightweight Open-Source Hallucination Detection System7 May 2025 0 repositories listed
-
Interpretable Zero-shot Learning with Infinite Class Concepts6 May 2025 0 repositories listed
-
Mitigating Image Captioning Hallucinations in Vision-Language Models6 May 2025 0 repositories listed
-
Knowledge Graphs for Enhancing Large Language Models in Entity Disambiguation5 May 2025 0 repositories listed
-
A Comprehensive Analysis for Visual Object Hallucination in Large Vision-Language Models4 May 2025 0 repositories listed
-
SEval-Ex: A Statement-Level Framework for Explainable Summarization Evaluation4 May 2025 0 repositories listed
-
Automated Parsing of Engineering Drawings for Structured Information Extraction Using a Fine-tuned Document Understanding Transformer2 May 2025 0 repositories listed
-
Multi-agents based User Values Mining for Recommendation2 May 2025 0 repositories listed
-
HalluMix: A Task-Agnostic, Multi-Domain Benchmark for Real-World Hallucination Detection1 May 2025 0 repositories listed
-
Triggering Hallucinations in LLMs: A Quantitative Study of Prompt-Induced Hallucination in Large Language Models1 May 2025 0 repositories listed
-
Black-Box Visual Prompt Engineering for Mitigating Object Hallucination in Large Vision Language Models30 Apr 2025 0 repositories listed
-
Efficient and robust 3D blind harmonization for large domain gaps30 Apr 2025 0 repositories listed
-
Localizing Before Answering: A Hallucination Evaluation Benchmark for Grounded Medical Multimodal LLMs30 Apr 2025 0 repositories listed
-
MAC-Tuning: LLM Multi-Compositional Problem Reasoning with Enhanced Knowledge Boundary Awareness30 Apr 2025 0 repositories listed
-
Can LLMs Detect Intrinsic Hallucinations in Paraphrasing and Machine Translation?29 Apr 2025 0 repositories listed
-
Hallucination by Code Generation LLMs: Taxonomy, Benchmarks, Mitigation, and Challenges29 Apr 2025 0 repositories listed
-
An Automated Reinforcement Learning Reward Design Framework with Large Language Model for Cooperative Platoon Coordination28 Apr 2025 0 repositories listed
-
Explanatory Summarization with Discourse-Driven Planning27 Apr 2025 0 repositories listed
-
Validating Network Protocol Parsers with Traceable RFC Document Interpretation25 Apr 2025 0 repositories listed
-
Data-Driven Calibration of Prediction Sets in Large Vision-Language Models Based on Inductive Conformal Prediction24 Apr 2025 0 repositories listed
-
Toward Personalizing Quantum Computing Education: An Evolutionary LLM-Powered Approach24 Apr 2025 0 repositories listed
-
(Im)possibility of Automated Hallucination Detection in Large Language Models23 Apr 2025 0 repositories listed
-
The Dance of Atoms-De Novo Protein Design with Diffusion Model23 Apr 2025 0 repositories listed
-
Grounded in Context: Retrieval-Based Method for Hallucination Detection22 Apr 2025 0 repositories listed
-
Insights from Verification: Training a Verilog Generation LLM with Reinforcement Learning with Testbench Feedback22 Apr 2025 0 repositories listed
-
aiXamine: Simplified LLM Safety and Security21 Apr 2025 0 repositories listed
-
POLYRAG: Integrating Polyviews into Retrieval-Augmented Generation for Medical Applications21 Apr 2025 0 repositories listed
-
ResNetVLLM-2: Addressing ResNetVLLM's Multi-Modal Hallucinations20 Apr 2025 0 repositories listed
-
Density Measures for Language Generation19 Apr 2025 0 repositories listed
-
Hydra: An Agentic Reasoning Approach for Enhancing Adversarial Robustness and Mitigating Hallucinations in Vision-Language Models19 Apr 2025 0 repositories listed
-
Multi-Stage Retrieval for Operational Technology Cybersecurity Compliance Using Large Language Models: A Railway Casestudy18 Apr 2025 0 repositories listed
-
Aspect-Based Summarization with Self-Aspect Retrieval Enhanced Generation17 Apr 2025 0 repositories listed
-
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training17 Apr 2025 0 repositories listed
-
QLLM: Do We Really Need a Mixing Network for Credit Assignment in Multi-Agent Reinforcement Learning?17 Apr 2025 0 repositories listed
-
Efficient Contrastive Decoding with Probabilistic Hallucination Detection - Mitigating Hallucinations in Large Vision Language Models -16 Apr 2025 0 repositories listed
-
Naming is framing: How cybersecurity's language problems are repeating in AI governance16 Apr 2025 0 repositories listed
-
Purposefully Induced Psychosis (PIP): Embracing Hallucination as Imagination in Large Language Models16 Apr 2025 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.