HalluScope: Fine-grained Hallucination Diagnosis for Multimodal Large Language Models
Merged summary
TL;DR - HalluScope introduces a unified framework for detecting, classifying, and explaining hallucinations in multimodal large language models. Its fine-grained diagnoses can also help other models correct hallucinated outputs.
- HalluScope-30K covers eight hallucination sources and five task categories.
- HalluScope-4B and HalluScope-8B use a multi-granular joint reward to optimize detection and classification together.
- The models achieve state-of-the-art results on MHALO and a fine-grained hallucination classification benchmark.
- Diagnosis-driven feedback improves hallucination correction in Qwen3-VL-8B-Instruct and LLaVA-1.5-7B.
Sources (1)
HalluScope: Fine-grained Hallucination Diagnosis for Multimodal Large Language Models
TL;DR - HalluScope introduces a unified framework for detecting, classifying, and explaining hallucinations in multimodal large language models. Its fine-grained diagnoses can also help other models correct hallucinated outputs.
- HalluScope-30K covers eight hallucination sources and five task categories.
- HalluScope-4B and HalluScope-8B use a multi-granular joint reward to optimize detection and classification together.
- The models achieve state-of-the-art results on MHALO and a fine-grained hallucination classification benchmark.
- Diagnosis-driven feedback improves hallucination correction in Qwen3-VL-8B-Instruct and LLaVA-1.5-7B.