Browse State-of-the-Art › Hallucination › Papers, page 13
Hallucination
Papers archive 2025-07-28
archive papers tagged: 1,816 · with a code link: 752 · where Syntology ran a sample: 276 (240 with a run with no instrument failure, 36 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (276 of 1,816 tagged: 240 with a run with no instrument failure, 36 where every run was a failure of Syntology's instrument)
Page 13 of 19: papers 1,201 to 1,300 of 1,816, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
ReDeEP: Detecting Hallucination in Retrieval-Augmented Generation via Mechanistic Interpretability15 Oct 2024 0 repositories listed
-
Can Structured Data Reduce Epistemic Uncertainty?14 Oct 2024 0 repositories listed
-
Medico: Towards Hallucination Detection and Correction with Multi-source Evidence Fusion14 Oct 2024 0 repositories listed
-
Parenting: Optimizing Knowledge Selection of Retrieval-Augmented Language Models with Parameter Decoupling and Tailored Tuning14 Oct 2024 0 repositories listed
-
SkillAggregation: Reference-free LLM-Dependent Aggregation14 Oct 2024 0 repositories listed
-
Collu-Bench: A Benchmark for Predicting Language Model Hallucinations in Code13 Oct 2024 0 repositories listed
-
Honest AI: Fine-Tuning "Small" Language Models to Say "I Don't Know", and Reducing Hallucination in RAG13 Oct 2024 0 repositories listed
-
12 Oct 2024 0 repositories listed Syntology 5 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples) · 1 pointer-only (licence)
-
Measuring the Inconsistency of Large Language Models in Preferential Ranking11 Oct 2024 0 repositories listed
-
Can Knowledge Graphs Make Large Language Models More Trustworthy? An Empirical Study over Open-ended Question Answering10 Oct 2024 0 repositories listed
-
LatteCLIP: Unsupervised CLIP Fine-Tuning via LMM-Synthetic Texts10 Oct 2024 0 repositories listed
-
PublicHearingBR: A Brazilian Portuguese Dataset of Public Hearing Transcripts for Summarization of Long Documents10 Oct 2024 0 repositories listed
-
From Pixels to Tokens: Revisiting Object Hallucinations in Large Vision-Language Models9 Oct 2024 0 repositories listed
-
EMMA: Empowering Multi-modal Mamba with Structural and Hierarchical Alignment8 Oct 2024 0 repositories listed
-
FG-PRM: Fine-grained Hallucination Detection and Mitigation in Language Model Mathematical Reasoning8 Oct 2024 0 repositories listed
-
Gradual Learning: Optimizing Fine-Tuning with Partially Mastered Knowledge in Large Language Models8 Oct 2024 0 repositories listed
-
Listening to Patients: A Framework of Detecting and Mitigating Patient Misreport for Medical Dialogue Generation8 Oct 2024 0 repositories listed
-
AI-Enhanced Ethical Hacking: A Linux-Focused Experiment7 Oct 2024 0 repositories listed
-
TLDR: Token-Level Detective Reward Model for Large Vision Language Models7 Oct 2024 0 repositories listed
-
DAMRO: Dive into the Attention Mechanism of LVLM to Reduce Object Hallucination6 Oct 2024 0 repositories listed
-
Mitigating Hallucinations Using Ensemble of Knowledge Graph and Vector Store in Large Language Models to Enhance Mental Health Support6 Oct 2024 0 repositories listed
-
DiDOTS: Knowledge Distillation from Large-Language-Models for Dementia Obfuscation in Transcribed Speech5 Oct 2024 0 repositories listed
-
Auto-GDA: Automatic Domain Adaptation for Efficient Grounding Verification in Retrieval Augmented Generation4 Oct 2024 0 repositories listed
-
SAG: Style-Aligned Article Generation via Model Collaboration4 Oct 2024 0 repositories listed
-
FactCheckmate: Preemptively Detecting and Mitigating Hallucinations in LMs3 Oct 2024 0 repositories listed
-
Enhancing Training Data Attribution for Large Language Models with Fitting Error Consideration2 Oct 2024 0 repositories listed
-
LMOD: A Large Multimodal Ophthalmology Dataset and Benchmark for Large Vision-Language Models2 Oct 2024 0 repositories listed
-
The Labyrinth of Links: Navigating the Associative Maze of Multi-modal LLMs2 Oct 2024 0 repositories listed
-
VideoCLIP-XL: Advancing Long Description Understanding for Video CLIP Models1 Oct 2024 0 repositories listed
-
Contrastive Token Learning with Similarity Decay for Repetition Suppression in Machine Translation30 Sep 2024 0 repositories listed
-
Ingest-And-Ground: Dispelling Hallucinations from Continually-Pretrained LLMs with RAG30 Sep 2024 0 repositories listed
-
MedHalu: Hallucinations in Responses to Healthcare Queries by Large Language Models29 Sep 2024 0 repositories listed
-
DENEB: A Hallucination-Robust Automatic Evaluation Metric for Image Captioning28 Sep 2024 0 repositories listed
-
26 Sep 2024 0 repositories listed Syntology 14 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 0 violated, 9 with no contract checked; 4 where Syntology's instrument failed) · 2 unverified (of 16 harvested samples) · 16 pointer-only (licence)
-
Enhancing Guardrails for Safe and Secure Healthcare AI25 Sep 2024 0 repositories listed
-
RoleBreak: Character Hallucination as a Jailbreak Attack in Role-Playing Systems25 Sep 2024 0 repositories listed
-
AsthmaBot: Multi-modal, Multi-Lingual Retrieval Augmented Generation For Asthma Patient Support24 Sep 2024 0 repositories listed
-
Enhancing Text-to-SQL Capabilities of Large Language Models via Domain Database Knowledge Injection24 Sep 2024 0 repositories listed
-
Planning in the Dark: LLM-Symbolic Planning Pipeline without Experts24 Sep 2024 0 repositories listed
-
Long-horizon Embodied Planning with Implicit Logical Inference and Hallucination Mitigation24 Sep 2024 0 repositories listed
-
A Preliminary Study of o1 in Medicine: Are We Closer to an AI Doctor?23 Sep 2024 0 repositories listed
-
Enhancing Scientific Reproducibility Through Automated BioCompute Object Creation Using Retrieval-Augmented Generation from Publications23 Sep 2024 0 repositories listed
-
Effectively Enhancing Vision Language Large Models by Prompt Augmentation and Caption Utilization22 Sep 2024 0 repositories listed
-
Contrastive Learning for Knowledge-Based Question Generation in Large Language Models21 Sep 2024 0 repositories listed
-
FIHA: Autonomous Hallucination Evaluation in Vision-Language Models with Davidson Scene Graphs20 Sep 2024 0 repositories listed
-
A Multiple-Fill-in-the-Blank Exam Approach for Enhancing Zero-Resource Hallucination Detection in Large Language Models20 Sep 2024 0 repositories listed
-
LLMs Can Check Their Own Results to Mitigate Hallucinations in Traffic Understanding Tasks19 Sep 2024 0 repositories listed
-
Textualized Agent-Style Reasoning for Complex Tasks by Multiple Round LLM Generation19 Sep 2024 0 repositories listed
-
Depth-based Privileged Information for Boosting 3D Human Pose Estimation on RGB17 Sep 2024 0 repositories listed
-
Zero-resource Hallucination Detection for Text Generation via Graph-based Contextual Knowledge Triples Modeling17 Sep 2024 0 repositories listed
-
Optimizing Resource Consumption in Diffusion Models through Hallucination Early Detection16 Sep 2024 0 repositories listed
-
SFR-RAG: Towards Contextually Faithful LLMs16 Sep 2024 0 repositories listed
-
Explore the Hallucination on Low-level Perception for MLLMs15 Sep 2024 0 repositories listed
-
ODE: Open-Set Evaluation of Hallucinations in Multimodal Large Language Models14 Sep 2024 0 repositories listed
-
Winning Solution For Meta KDD Cup' 2413 Sep 2024 0 repositories listed
-
MEDIC: Towards a Comprehensive Framework for Evaluating LLMs in Clinical Applications11 Sep 2024 0 repositories listed
-
Safety challenges of AI in medicine in the era of large language models11 Sep 2024 0 repositories listed
-
Mitigating Hallucination in Visual-Language Models via Re-Balancing Contrastive Decoding10 Sep 2024 0 repositories listed
-
LLMs Will Always Hallucinate, and We Need to Live With This9 Sep 2024 0 repositories listed
-
Combining LLMs and Knowledge Graphs to Reduce Hallucinations in Question Answering6 Sep 2024 0 repositories listed
-
Detecting Buggy Contracts via Smart Testing6 Sep 2024 0 repositories listed
-
Vietnamese Legal Information Retrieval in Question-Answering System5 Sep 2024 0 repositories listed
-
CLUE: Concept-Level Uncertainty Estimation for Large Language Models4 Sep 2024 0 repositories listed
-
Improved Single Camera BEV Perception Using Multi-Camera Training4 Sep 2024 0 repositories listed
-
What does it take to get state of the art in simultaneous speech-to-speech translation?2 Sep 2024 0 repositories listed
-
LLMs Prompted for Graphs: Hallucinations and Generative Capabilities30 Aug 2024 0 repositories listed
-
Pre-Training Multimodal Hallucination Detectors with Corrupted Grounding Data30 Aug 2024 0 repositories listed
-
UserSumBench: A Benchmark Framework for Evaluating User Summarization Approaches30 Aug 2024 0 repositories listed
-
Evidence-Enhanced Triplet Generation Framework for Hallucination Alleviation in Generative Question Answering27 Aug 2024 0 repositories listed
-
Measuring text summarization factuality using atomic facts entailment metrics in the context of retrieval augmented generation27 Aug 2024 0 repositories listed
-
Negation Blindness in Large Language Models: Unveiling the NO Syndrome in Image Generation27 Aug 2024 0 repositories listed
-
Towards Reliable Medical Question Answering: Techniques and Challenges in Mitigating Hallucinations in Language Models25 Aug 2024 0 repositories listed
-
Can LLM be a Good Path Planner based on Prompt Engineering? Mitigating the Hallucination for Path Planning23 Aug 2024 0 repositories listed
-
Internal and External Knowledge Interactive Refinement Framework for Knowledge-Intensive Question Answering23 Aug 2024 0 repositories listed
-
MedDiT: A Knowledge-Controlled Diffusion Transformer Framework for Dynamic Medical Image Generation in Virtual Simulated Patient22 Aug 2024 0 repositories listed
-
RAG-Optimized Tibetan Tourism LLMs: Enhancing Accuracy and Personalization21 Aug 2024 0 repositories listed
-
Towards Analyzing and Mitigating Sycophancy in Large Vision-Language Models21 Aug 2024 0 repositories listed
-
CLIP-DPO: Vision-Language Models as a Source of Preference for Fixing Hallucinations in LVLMs19 Aug 2024 0 repositories listed
-
Enhanced document retrieval with topic embeddings19 Aug 2024 0 repositories listed
-
MAPLE: Enhancing Review Generation with Multi-Aspect Prompt LEarning in Explainable Recommendation19 Aug 2024 0 repositories listed
-
Cognitive LLMs: Towards Integrating Cognitive Architectures and Large Language Models for Manufacturing Decision-making17 Aug 2024 0 repositories listed
-
Large Language Models Might Not Care What You Are Saying: Prompt Format Beats Descriptions16 Aug 2024 0 repositories listed
-
Lower Layer Matters: Alleviating Hallucination via Multi-Layer Fusion Contrastive Decoding with Truthfulness Refocused16 Aug 2024 0 repositories listed
-
Plan with Code: Comparing approaches for robust NL to DSL generation15 Aug 2024 0 repositories listed
-
CodeMirage: Hallucinations in Code Generated by Large Language Models14 Aug 2024 0 repositories listed
-
Training Language Models on the Knowledge Graph: Insights on Hallucinations and Their Detectability14 Aug 2024 0 repositories listed
-
Audit-LLM: Multi-Agent Collaboration for Log-based Insider Threat Detection12 Aug 2024 0 repositories listed
-
Reference-free Hallucination Detection for Large Vision-Language Models11 Aug 2024 0 repositories listed
-
Improving Whisper's Recognition Performance for Under-Represented Language Kazakh Leveraging Unpaired Speech and Text10 Aug 2024 0 repositories listed
-
FiSTECH: Financial Style Transfer to Enhance Creativity without Hallucinations in LLMs9 Aug 2024 0 repositories listed
-
KnowPO: Knowledge-aware Preference Optimization for Controllable Knowledge Selection in Retrieval-Augmented Language Models6 Aug 2024 0 repositories listed
-
MAO: A Framework for Process Model Generation with Multi-Agent Orchestration4 Aug 2024 0 repositories listed
-
Improving Zero-Shot ObjectNav with Generative Communication3 Aug 2024 0 repositories listed
-
Piculet: Specialized Models-Guided Hallucination Decrease for MultiModal Large Language Models2 Aug 2024 0 repositories listed
-
Misinforming LLMs: vulnerabilities, challenges and opportunities2 Aug 2024 0 repositories listed
-
Alleviating Hallucination in Large Vision-Language Models with Active Retrieval Augmentation1 Aug 2024 0 repositories listed
-
Prompting Medical Large Vision-Language Models to Diagnose Pathologies by Visual Question Answering31 Jul 2024 0 repositories listed
-
Cost-Effective Hallucination Detection for LLMs31 Jul 2024 0 repositories listed
-
Interpreting and Mitigating Hallucination in MLLMs through Multi-agent Debate30 Jul 2024 0 repositories listed
-
24 Jul 2024 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.