Browse State-of-the-Art › Hallucination › Papers, page 11
Hallucination
Papers archive 2025-07-28
archive papers tagged: 1,816 · with a code link: 752 · where Syntology ran a sample: 276 (240 with a run with no instrument failure, 36 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (276 of 1,816 tagged: 240 with a run with no instrument failure, 36 where every run was a failure of Syntology's instrument)
Page 11 of 19: papers 1,001 to 1,100 of 1,816, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
MedHallu: A Comprehensive Benchmark for Detecting Medical Hallucinations in Large Language Models20 Feb 2025 0 repositories listed
-
Verify when Uncertain: Beyond Self-Consistency in Black Box Hallucination Detection20 Feb 2025 0 repositories listed
-
Detecting LLM Fact-conflicting Hallucinations Enhanced by Temporal-logic-based Reasoning19 Feb 2025 0 repositories listed
-
OpenSearch-SQL: Enhancing Text-to-SQL with Dynamic Few-shot and Consistency Alignment19 Feb 2025 0 repositories listed
-
REFIND: Retrieval-Augmented Factuality Hallucination Detection in Large Language Models19 Feb 2025 0 repositories listed
-
What are Models Thinking about? Understanding Large Language Model Hallucinations "Psychology" through Model Inner State Analysis19 Feb 2025 0 repositories listed
-
CutPaste&Find: Efficient Multimodal Hallucination Detector with Visual-aid Knowledge Base18 Feb 2025 0 repositories listed
-
Lost in Transcription, Found in Distribution Shift: Demystifying Hallucination in Speech Foundation Models18 Feb 2025 0 repositories listed
-
Can Your Uncertainty Scores Detect Hallucinated Entity?17 Feb 2025 0 repositories listed
-
A Survey of LLM-based Agents in Medicine: How far are we from Baymax?16 Feb 2025 0 repositories listed
-
Smoothing Out Hallucinations: Mitigating LLM Hallucination with Smoothed Knowledge Distillation16 Feb 2025 0 repositories listed
-
Valuable Hallucinations: Realizable Non-realistic Propositions16 Feb 2025 0 repositories listed
-
Enhancing RAG with Active Learning on Conversation Records: Reject Incapables and Answer Capables13 Feb 2025 0 repositories listed
-
DeepSeek on a Trip: Inducing Targeted Visual Hallucinations via Representation Vulnerabilities11 Feb 2025 0 repositories listed
-
Refine Knowledge of Large Language Models via Adaptive Contrastive Learning11 Feb 2025 0 repositories listed
-
Hallucination Detection: A Probabilistic Framework Using Embeddings Distance Analysis10 Feb 2025 0 repositories listed
-
ChallengeMe: An Adversarial Learning-enabled Text Summarization Framework7 Feb 2025 0 repositories listed
-
Enhancing Hallucination Detection through Noise Injection6 Feb 2025 0 repositories listed
-
TruthFlow: Truthful LLM Generation via Representation Flow Correction6 Feb 2025 0 repositories listed
-
A Schema-Guided Reason-while-Retrieve framework for Reasoning on Scene Graphs with Large-Language-Models (LLMs)5 Feb 2025 0 repositories listed
-
Mitigating Object Hallucinations in Large Vision-Language Models via Attention Calibration4 Feb 2025 0 repositories listed
-
Assessing the use of Diffusion models for motion artifact correction in brain MRI3 Feb 2025 0 repositories listed
-
Eliciting Language Model Behaviors with Investigator Agents3 Feb 2025 0 repositories listed
-
MJ-VIDEO: Fine-Grained Benchmarking and Rewarding Video Preferences in Video Generation3 Feb 2025 0 repositories listed
-
SelfCheckAgent: Zero-Resource Hallucination Detection in Generative Large Language Models3 Feb 2025 0 repositories listed
-
MINT: Mitigating Hallucinations in Large Vision-Language Models via Token Reduction2 Feb 2025 0 repositories listed
-
Importing Phantoms: Measuring LLM Package Hallucination Vulnerabilities31 Jan 2025 0 repositories listed
-
Poison as Cure: Visual Noise for Mitigating Object Hallucinations in LVMs31 Jan 2025 0 repositories listed
-
Few-Shot Optimized Framework for Hallucination Detection in Resource-Limited NLP Systems28 Jan 2025 0 repositories listed
-
Mitigating Hallucinated Translations in Large Language Models with Hallucination-focused Preference Optimization28 Jan 2025 0 repositories listed
-
Open-Source Retrieval Augmented Generation Framework for Retrieving Accurate Medication Insights from Formularies for African Healthcare Workers28 Jan 2025 0 repositories listed
-
Scaling Large Vision-Language Models for Enhanced Multimodal Comprehension In Biomedical Image Analysis26 Jan 2025 0 repositories listed
-
Evaluating Hallucination in Large Vision-Language Models based on Context-Aware Object Similarities25 Jan 2025 0 repositories listed
-
Mirage in the Eyes: Hallucination Attack on Multi-modal Large Language Models with Only Attention Sink25 Jan 2025 0 repositories listed
-
Measuring and Mitigating Hallucinations in Vision-Language Dataset Generation for Remote Sensing24 Jan 2025 0 repositories listed
-
Comprehensive Modeling and Question Answering of Cancer Clinical Practice Guidelines using LLMs23 Jan 2025 0 repositories listed
-
Hallucinations Can Improve Large Language Models in Drug Discovery23 Jan 2025 0 repositories listed
-
RAG-Reward: Optimizing RAG with Reward Modeling and RLHF22 Jan 2025 0 repositories listed
-
Question-to-Question Retrieval for Hallucination-Free Knowledge Access: An Approach for Wikipedia and Wikidata Question Answering20 Jan 2025 0 repositories listed
-
ArxEval: Evaluating Retrieval and Generation in Language Models for Scientific Literature17 Jan 2025 0 repositories listed
-
Attention-guided Self-reflection for Zero-shot Hallucination Detection in Large Language Models17 Jan 2025 0 repositories listed
-
FRAG: A Flexible Modular Framework for Retrieval-Augmented Generation based on Knowledge Graphs17 Jan 2025 0 repositories listed
-
A Survey on Responsible LLMs: Inherent Risk, Malicious Use, and Mitigation Strategy16 Jan 2025 0 repositories listed
-
HALoGEN: Fantastic LLM Hallucinations and Where to Find Them14 Jan 2025 0 repositories listed
-
GPT as a Monte Carlo Language Tree: A Probabilistic Perspective13 Jan 2025 0 repositories listed
-
MedCT: A Clinical Terminology Graph for Generative AI Applications in Healthcare11 Jan 2025 0 repositories listed
-
Hermit Kingdom Through the Lens of Multiple Perspectives: A Case Study of LLM Hallucination on North Korea10 Jan 2025 0 repositories listed
-
Seeing with Partial Certainty: Conformal Prediction for Robotic Scene Recognition in Built Environments9 Jan 2025 0 repositories listed
-
Feedback-Driven Vision-Language Alignment with Minimal Human Supervision8 Jan 2025 0 repositories listed
-
RAG-Check: Evaluating Multimodal Retrieval Augmented Generation Performance7 Jan 2025 0 repositories listed
-
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models6 Jan 2025 0 repositories listed
-
FlippedRAG: Black-Box Opinion Manipulation Adversarial Attacks to Retrieval-Augmented Generation Models6 Jan 2025 0 repositories listed
-
Foundations of GenIR6 Jan 2025 0 repositories listed
-
CarbonChat: Large Language Model-Based Corporate Carbon Emission Analysis and Climate Knowledge Q&A System3 Jan 2025 0 repositories listed
-
LLMs & Legal Aid: Understanding Legal Needs Exhibited Through User Queries3 Jan 2025 0 repositories listed
-
Enhancing Uncertainty Modeling with Semantic Graph for Hallucination Detection2 Jan 2025 0 repositories listed
-
Large Language Model-Enhanced Symbolic Reasoning for Knowledge Base Completion2 Jan 2025 0 repositories listed
-
IllusionBench: A Large-scale and Comprehensive Benchmark for Visual Illusion Understanding in Vision-Language Models1 Jan 2025 0 repositories listed
-
POPEN: Preference-Based Optimization and Ensemble for LVLM-Based Reasoning Segmentation1 Jan 2025 0 repositories listed
-
Stop Learning it all to Mitigate Visual Hallucination, Focus on the Hallucination Target.1 Jan 2025 0 repositories listed
-
VL-RewardBench: A Challenging Benchmark for Vision-Language Generative Reward Models1 Jan 2025 0 repositories listed
-
A review of faithfulness metrics for hallucination assessment in Large Language Models31 Dec 2024 0 repositories listed
-
Distilling Desired Comments for Enhanced Code Review with Large Language Models29 Dec 2024 0 repositories listed
-
Is Your Text-to-Image Model Robust to Caption Noise?27 Dec 2024 0 repositories listed
-
An End-to-End Depth-Based Pipeline for Selfie Image Rectification26 Dec 2024 0 repositories listed
-
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models25 Dec 2024 0 repositories listed
-
From Hallucinations to Facts: Enhancing Language Models with Curated Knowledge Graphs24 Dec 2024 0 repositories listed
-
Improving Factuality with Explicit Working Memory24 Dec 2024 0 repositories listed
-
AlzheimerRAG: Multimodal Retrieval Augmented Generation for PubMed articles21 Dec 2024 0 repositories listed
-
Logical Consistency of Large Language Models in Fact-checking20 Dec 2024 0 repositories listed
-
Toward Robust Hyper-Detailed Image Captioning: A Multiagent Approach and Dual Evaluation Metrics for Factuality and Coverage20 Dec 2024 0 repositories listed
-
A Comparative Study of DSPy Teleprompter Algorithms for Aligning Large Language Models Evaluation Metrics to Human Evaluation19 Dec 2024 0 repositories listed
-
Dehallucinating Parallel Context Extension for Retrieval-Augmented Generation19 Dec 2024 0 repositories listed
-
Query pipeline optimization for cancer patient question answering systems19 Dec 2024 0 repositories listed
-
Think&Cite: Improving Attributed Text Generation with Self-Guided Tree Search and Progress Reward Modeling19 Dec 2024 0 repositories listed
-
Token Preference Optimization with Self-Calibrated Visual-Anchored Rewards for Hallucination Mitigation19 Dec 2024 0 repositories listed
-
Are LLMs Good Literature Review Writers? Evaluating the Literature Review Writing Ability of Large Language Models18 Dec 2024 0 repositories listed
-
Cracking the Code of Hallucination in LVLMs with Vision-aware Head Divergence18 Dec 2024 0 repositories listed
-
A MapReduce Approach to Effectively Utilize Long Context Information in Retrieval Augmented Language Models17 Dec 2024 0 repositories listed
-
ReXTrust: A Model for Fine-Grained Hallucination Detection in AI-Generated Radiology Reports17 Dec 2024 0 repositories listed
-
What External Knowledge is Preferred by LLMs? Characterizing and Exploring Chain of Evidence in Imperfect Context17 Dec 2024 0 repositories listed
-
When to Speak, When to Abstain: Contrastive Decoding with Abstention17 Dec 2024 0 repositories listed
-
CG-Bench: Clue-grounded Question Answering Benchmark for Long Video Understanding16 Dec 2024 0 repositories listed
-
Combating Multimodal LLM Hallucination via Bottom-Up Holistic Reasoning15 Dec 2024 0 repositories listed
-
RAC3: Retrieval-Augmented Corner Case Comprehension for Autonomous Driving with Vision-Language Models15 Dec 2024 0 repositories listed
-
Task-Oriented Dialog Systems for the Senegalese Wolof Language15 Dec 2024 0 repositories listed
-
Accelerating Retrieval-Augmented Generation14 Dec 2024 0 repositories listed
-
NoisyEQA: Benchmarking Embodied Question Answering Against Noisy Queries14 Dec 2024 0 repositories listed
-
Thinking with Knowledge Graphs: Enhancing LLM Reasoning Through Structured Data14 Dec 2024 0 repositories listed
-
Benchmarking large language models for materials synthesis: the case of atomic layer deposition13 Dec 2024 0 repositories listed
-
Detecting LLM Hallucination Through Layer-wise Information Deficiency: Analysis of Unanswerable Questions and Ambiguous Prompts13 Dec 2024 0 repositories listed
-
TACOMORE: Leveraging the Potential of LLMs in Corpus-based Discourse Analysis with Prompt Engineering13 Dec 2024 0 repositories listed
-
12 Dec 2024 0 repositories listed
-
HalluCana: Fixing LLM Hallucination with A Canary Lookahead10 Dec 2024 0 repositories listed
-
Methods for Legal Citation Prediction in the Age of LLMs: An Australian Law Case Study9 Dec 2024 0 repositories listed
-
Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent7 Dec 2024 0 repositories listed
-
100% Elimination of Hallucinations on RAGTruth for GPT-4 and GPT-3.5 Turbo6 Dec 2024 0 repositories listed
-
TOBUGraph: Knowledge Graph-Based Retrieval for Enhanced LLM Performance Beyond RAG6 Dec 2024 0 repositories listed
-
LLM-Align: Utilizing Large Language Models for Entity Alignment in Knowledge Graphs6 Dec 2024 0 repositories listed
-
Multi-Objective Alignment of Large Language Models Through Hypervolume Maximization6 Dec 2024 0 repositories listed