Skip to content
Preprint

Does Playing it Safe Count as Faithfulness? Reassessing LVLM Hallucination Mitigation Methods

Sep 2026 · 0 citations · 25 references
Computer Science

TL;DR

It is argued that hallucination mitigation should be evaluated as a faithfulness--informativeness--capability trade-off rather than through hallucination scores alone, because improvements on hallucination benchmarks do not reliably transfer to broader multimodal capabilities.

Abstract

Recent inference-time hallucination mitigation methods for large vision-language models (LVLMs) report strong gains on hallucination benchmarks. However, it remains unclear whether lower hallucination scores reflect improved multimodal grounding or more conservative generation. We evaluate six mitigation methods across three LVLMs and four benchmarks, including hallucination-focused evaluation and the diverse capability benchmark MMStar. Our analysis reveals two consistent patterns. First, hallucination reduction is often coupled with reduced informativeness: methods that lower hallucination rates also reduce object recall, visual coverage, or response detailedness. Second, improvements on hallucination benchmarks do not reliably transfer to broader multimodal capabilities, with methods showing inconsistent or degraded performance on fine-grained perception and reasoning tasks. Our findings suggest that current evaluation protocols may overestimate progress by rewarding conservative generation. We argue that hallucination mitigation should be evaluated as a faithfulness--informativeness--capability trade-off rather than through hallucination scores alone.

View source

Similar papers

#artificial intelligence Preprint Sep 2026

Beneath the Scores: Rethinking Hallucination Evaluation for Video Understanding Models

Video understanding is increasingly performed by multi-stage LLM agents that separate temporal grounding, visual observation, and reasoning. Yet these stages are typically evaluated on different benchmarks and distributions, making it difficult to determine where hallucinations originate. We first organize existing ben...

Shu-Zhi Gong, F. Sun, Yuansan Liu · 0 citations
Preprint Sep 2026

What Do Hallucinations Reveal About Multimodal Reasoning? Diagnosing Visual Grounding Failures via Contrastive Decoding Probes

When strong multimodal models are widely available, progress requires new scientific methodologies beyond benchmark scores---using models as instruments for understanding behavior. We address this by asking: can we use large vision-language models (LVLMs) as experimental instruments for studying their own failure dynam...

Zhi-Peng Zhao, Wen-Xu Wang, Peishun Liu et al. · 0 citations
#machine learning Preprint Sep 2026

MISHAP-Bench: A Hallucination Benchmark for Large Audio-Language Models

This work introduces MISHAP-Bench, a comprehensive benchmark with 12,000 challenging open-ended question-audio pairs and a rigorous evaluation pipeline covering two hallucination categories, and proposes a groundedness judge that uses reference rubrics and judge prompts guided by human annotations.

Wen-Soi Zhi, Giulio Segalini, Jian-Jia Chen et al. · 0 citations
Book Open access Aug 2026

Who's Adam? Benchmarking Hallucinations in Scientific Dialogue

ADAM-Bench (Auditing Dialogue Assertions with Multimodal Evidence), a benchmark for paper-grounded hallucinations in scientific dialogue, is introduced and two tasks are defined: hallucination detection and minimal evidence set localization.

Ze-Xing Zhang, Tian-Yang Lei, Ke-Wei Yang et al. · 0 citations
Preprint Aug 2026

UHP Detection: LVLMs have their Unique Hallucination Pattern in the Consistency Space

Large vision--language models (LVLMs) demonstrate strong multimodal reasoning capabilities but remain prone to hallucination, where model predictions are not grounded in visual evidence, so a fully black-box framework that models hallucination as a structured uncertainty pattern is proposed.

Amir Mohammad Ezzati, Kiyan Rezaee, Bardiya Kariminia et al. · 0 citations
Preprint Sep 2026

Multi-Faceted Evaluation and Mitigation of Emotion Hallucinations in MLLMs

Multimodal large language models (MLLMs) have shown strong potential in open-ended emotion understanding, yet they often generate emotion hallucinations. Evaluating such hallucinations is particularly challenging for two reasons. First, emotion understanding spans multiple cognitive facets, from multimodal perception t...

Bo-Wen Zeng, Pei-Pei Song, Wei-Dong Chen et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.