Browse State-of-the-Art › Hallucination › Papers, page 10
Hallucination
Papers archive 2025-07-28
archive papers tagged: 1,816 · with a code link: 752 · where Syntology ran a sample: 276 (240 with a run with no instrument failure, 36 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (276 of 1,816 tagged: 240 with a run with no instrument failure, 36 where every run was a failure of Syntology's instrument)
Page 10 of 19: papers 901 to 1,000 of 1,816, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
16 Apr 2025 0 repositories listed
-
From Misleading Queries to Accurate Answers: A Three-Stage Fine-Tuning Method for LLMs15 Apr 2025 0 repositories listed
-
Hallucination-Aware Generative Pretrained Transformer for Cooperative Aerial Mobility Control15 Apr 2025 0 repositories listed
-
Hallucination Detection in LLMs via Topological Divergence on Attention Graphs14 Apr 2025 0 repositories listed
-
The Future of MLLM Prompting is Adaptive: A Comprehensive Experimental Evaluation of Prompt Engineering Methods for Robust Multimodal Performance14 Apr 2025 0 repositories listed
-
DiTSE: High-Fidelity Generative Speech Enhancement via Latent Diffusion Transformers13 Apr 2025 0 repositories listed
-
Enhancing Mathematical Reasoning in Large Language Models with Self-Consistency-Based Hallucination Detection13 Apr 2025 0 repositories listed
-
SynthTRIPs: A Knowledge-Grounded Framework for Benchmark Query Generation for Personalized Tourism Recommenders12 Apr 2025 0 repositories listed
-
Cross-Document Cross-Lingual NLI via RST-Enhanced Graph Fusion and Interpretability Prediction11 Apr 2025 0 repositories listed
-
Hallucination, reliability, and the role of generative AI in science11 Apr 2025 0 repositories listed
-
MedHal: An Evaluation Dataset for Medical Hallucination Detection11 Apr 2025 0 repositories listed
-
Generative AI in Collaborative Academic Report Writing: Advantages, Disadvantages, and Ethical Considerations10 Apr 2025 0 repositories listed
-
How to Detect and Defeat Molecular Mirage: A Metric-Driven Benchmark for Hallucination in LLM-based Molecular Comprehension10 Apr 2025 0 repositories listed
-
10 Apr 2025 0 repositories listed Syntology 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Endowing Embodied Agents with Spatial Reasoning Capabilities for Vision-and-Language Navigation9 Apr 2025 0 repositories listed
-
OLMoTrace: Tracing Language Model Outputs Back to Trillions of Training Tokens9 Apr 2025 0 repositories listed
-
Perception in Reflection9 Apr 2025 0 repositories listed
-
Graph-based Approaches and Functionalities in Retrieval-Augmented Generation: A Comprehensive Survey8 Apr 2025 0 repositories listed
-
Capturing AI's Attention: Physics of Repetition, Hallucination, Bias and Beyond6 Apr 2025 0 repositories listed
-
TARAC: Mitigating Hallucination in LVLMs via Temporal Attention Real-time Accumulative Connection5 Apr 2025 0 repositories listed
-
Bridging LMS and Generative AI: Dynamic Course Content Integration (DCCI) for Connecting LLMs to Course Content -- The Ask ME Assistant4 Apr 2025 0 repositories listed
-
Hallucination Detection on a Budget: Efficient Bayesian Estimation of Semantic Entropy4 Apr 2025 0 repositories listed
-
Practical Poisoning Attacks against Retrieval-Augmented Generation4 Apr 2025 0 repositories listed
-
A Memory-Augmented LLM-Driven Method for Autonomous Merging of 3D Printing Work Orders3 Apr 2025 0 repositories listed
-
A Unified Virtual Mixture-of-Experts Framework:Enhanced Inference and Hallucination Mitigation in Single-Model System1 Apr 2025 0 repositories listed
-
GraphMaster: Automated Graph Synthesis via LLM Agents in Data-Limited Environments1 Apr 2025 0 repositories listed
-
The Illusionist's Prompt: Exposing the Factual Vulnerabilities of Large Language Models with Linguistic Nuances1 Apr 2025 0 repositories listed
-
An Analysis of Decoding Methods for LLM-based Agents for Faithful Multi-Hop Question Answering30 Mar 2025 0 repositories listed
-
Learning to Instruct for Visual Instruction Tuning28 Mar 2025 0 repositories listed
-
Alleviating LLM-based Generative Retrieval Hallucination in Alipay Search27 Mar 2025 0 repositories listed
-
Real-Time Evaluation Models for RAG: Who Detects Hallucinations Best?27 Mar 2025 0 repositories listed
-
Tricking Retrievers with Influential Tokens: An Efficient Black-Box Corpus Poisoning Attack27 Mar 2025 0 repositories listed
-
Instruction-Oriented Preference Alignment for Enhancing Multi-Modal Comprehension Capability of MLLMs26 Mar 2025 0 repositories listed
-
Mitigating Low-Level Visual Hallucinations Requires Self-Awareness: Database, Model and Training Strategy26 Mar 2025 0 repositories listed
-
Vision-Amplified Semantic Entropy for Hallucination Detection in Medical Visual Question Answering26 Mar 2025 0 repositories listed
-
HausaNLP at SemEval-2025 Task 3: Towards a Fine-Grained Model-Aware Hallucination Detection25 Mar 2025 0 repositories listed
-
KSHSeek: Data-Driven Approaches to Mitigating and Detecting Knowledge-Shortcut Hallucinations in Generative Models25 Mar 2025 0 repositories listed
-
ShED-HD: A Shannon Entropy Distribution Framework for Lightweight Hallucination Detection on Edge Devices23 Mar 2025 0 repositories listed
-
good4cir: Generating Detailed Synthetic Captions for Composed Image Retrieval22 Mar 2025 0 repositories listed
-
FactSelfCheck: Fact-Level Black-Box Hallucination Detection for LLMs21 Mar 2025 0 repositories listed
-
Judge Anything: MLLM as a Judge Across Any Modality21 Mar 2025 0 repositories listed
-
DNR Bench: Benchmarking Over-Reasoning in Reasoning LLMs20 Mar 2025 0 repositories listed
-
ECKGBench: Benchmarking Large Language Models in E-commerce Leveraging Knowledge Graph20 Mar 2025 0 repositories listed
-
MASH-VLM: Mitigating Action-Scene Hallucination in Video-LLMs through Disentangled Spatial-Temporal Representations20 Mar 2025 0 repositories listed
-
MMDT: Decoding the Trustworthiness and Safety of Multimodal Foundation Models19 Mar 2025 0 repositories listed
-
19 Mar 2025 0 repositories listed
-
R²: A LLM Based Novel-to-Screenplay Generation Framework with Causal Plot Graphs19 Mar 2025 0 repositories listed
-
Enhancing LLM Generation with Knowledge Hypergraph for Evidence-Based Medicine18 Mar 2025 0 repositories listed
-
From "Hallucination" to "Suture": Insights from Language Philosophy to Enhance Large Language Models18 Mar 2025 0 repositories listed
-
RAD: Retrieval-Augmented Decision-Making of Meta-Actions with Vision-Language Models in Autonomous Driving18 Mar 2025 0 repositories listed
-
LLMSeR: Enhancing Sequential Recommendation via LLM-based Data Augmentation16 Mar 2025 0 repositories listed
-
Applications of Large Language Model Reasoning in Feature Generation15 Mar 2025 0 repositories listed
-
LLM Agents for Education: Advances and Applications14 Mar 2025 0 repositories listed
-
RAG-KG-IL: A Multi-Agent Hybrid Framework for Reducing Hallucinations and Enhancing LLM Reasoning through RAG and Incremental Knowledge Graph Learning Integration14 Mar 2025 0 repositories listed
-
Learning to Inference Adaptively for Multimodal Large Language Models13 Mar 2025 0 repositories listed
-
Through the Magnifying Glass: Adaptive Perception Magnification for Hallucination-Free VLM Decoding13 Mar 2025 0 repositories listed
-
Is LLMs Hallucination Usable? LLM-based Negative Reasoning for Fake News Detection12 Mar 2025 0 repositories listed
-
Attention Hijackers: Detect and Disentangle Attention Hijacking in LVLMs for Hallucination Mitigation11 Mar 2025 0 repositories listed
-
Attention Reallocation: Towards Zero-cost and Controllable Hallucination Mitigation of MLLMs11 Mar 2025 0 repositories listed
-
Gradient-guided Attention Map Editing: Towards Efficient Contextual Hallucination Mitigation11 Mar 2025 0 repositories listed
-
OmniPaint: Mastering Object-Oriented Editing via Disentangled Insertion-Removal Inpainting11 Mar 2025 0 repositories listed
-
Seeing What's Not There: Spurious Correlation in Multimodal LLMs11 Mar 2025 0 repositories listed
-
Benchmarking Chinese Medical LLMs: A Medbench-based Analysis of Performance Gaps and Hierarchical Optimization Strategies10 Mar 2025 0 repositories listed
-
CtrlRAG: Black-box Adversarial Attacks Based on Masked Language Models in Retrieval-Augmented Language Generation10 Mar 2025 0 repositories listed
-
EAZY: Eliminating Hallucinations in LVLMs by Zeroing out Hallucinatory Image Tokens10 Mar 2025 0 repositories listed
-
Mitigating Hallucinations in YOLO-based Object Detection Models: A Revisit to Out-of-Distribution Detection10 Mar 2025 0 repositories listed
-
CalliReader: Contextualizing Chinese Calligraphy via an Embedding-Aligned Vision-Language Model9 Mar 2025 0 repositories listed
-
PerturboLLaVA: Reducing Multimodal Hallucinations with Perturbative Visual Training9 Mar 2025 0 repositories listed
-
Maximum Hallucination Standards for Domain-Specific Large Language Models7 Mar 2025 0 repositories listed
-
SINdex: Semantic INconsistency Index for Hallucination Detection in LLMs7 Mar 2025 0 repositories listed
-
LVLM-Compress-Bench: Benchmarking the Broader Impact of Large Vision-Language Model Compression6 Mar 2025 0 repositories listed
-
TPC: Cross-Temporal Prediction Connection for Vision-Language Model Hallucination Reduction6 Mar 2025 0 repositories listed
-
DSVD: Dynamic Self-Verify Decoding for Faithful Generation in Large Language Models5 Mar 2025 0 repositories listed
-
Monitoring Decoding: Mitigating Hallucination via Evaluating the Factuality of Partial Response during Generation5 Mar 2025 0 repositories listed
-
See What You Are Told: Visual Attention Sink in Large Multimodal Models5 Mar 2025 0 repositories listed
-
Towards Understanding Text Hallucination of Diffusion Models via Local Generation Bias5 Mar 2025 0 repositories listed
-
SAFE: A Sparse Autoencoder-Based Framework for Robust Query Enrichment and Hallucination Mitigation in LLMs4 Mar 2025 0 repositories listed
-
Tackling Hallucination from Conditional Models for Medical Image Reconstruction with DynamicDPS3 Mar 2025 0 repositories listed
-
LLM-Advisor: An LLM Benchmark for Cost-efficient Path Planning across Multiple Terrains3 Mar 2025 0 repositories listed
-
Adaptively profiling models with task elicitation3 Mar 2025 0 repositories listed
-
Explainable Depression Detection in Clinical Interviews with Personalized Retrieval-Augmented Generation3 Mar 2025 0 repositories listed
-
Unmasking Digital Falsehoods: A Comparative Analysis of LLM-Based Misinformation Detection Strategies2 Mar 2025 0 repositories listed
-
Steer LLM Latents for Hallucination Detection1 Mar 2025 0 repositories listed
-
1 Mar 2025 0 repositories listed
-
Semantic Volume: Quantifying and Detecting both External and Internal Uncertainty in LLMs28 Feb 2025 0 repositories listed
-
Exploring the Generalizability of Factual Hallucination Mitigation via Enhancing Precise Knowledge Utilization26 Feb 2025 0 repositories listed
-
On the Importance of Text Preprocessing for Multimodal Representation Learning and Pathology Report Generation26 Feb 2025 0 repositories listed
-
Winning Big with Small Models: Knowledge Distillation vs. Self-Training for Reducing Hallucination in QA Agents26 Feb 2025 0 repositories listed
-
BRIDO: Bringing Democratic Order to Abstractive Summarization25 Feb 2025 0 repositories listed
-
Exploring Causes and Mitigation of Hallucinations in Large Vision Language Models24 Feb 2025 0 repositories listed
-
`Generalization is hallucination' through the lens of tensor completions24 Feb 2025 0 repositories listed
-
The Law of Knowledge Overshadowing: Towards Understanding, Predicting, and Preventing LLM Hallucination22 Feb 2025 0 repositories listed
-
Uncertainty-Aware Fusion: An Ensemble Framework for Mitigating Hallucinations in Large Language Models22 Feb 2025 0 repositories listed
-
ZiGong 1.0: A Large Language Model for Financial Credit22 Feb 2025 0 repositories listed
-
The Role of Background Information in Reducing Object Hallucination in Vision-Language Models: Insights from Cutoff API Prompting21 Feb 2025 0 repositories listed
-
Hallucination Detection in Large Language Models with Metamorphic Relations20 Feb 2025 0 repositories listed
-
Large Language Models Struggle to Describe the Haystack without Human Help: Human-in-the-loop Evaluation of LLMs20 Feb 2025 0 repositories listed
Syntology lines on 4 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.