Do We Really Need External Tools to Mitigate Hallucinations? SIRA: Shared-Prefix Internal Reconstruction of Attribution

About

Large vision-language models (LVLMs) often hallucinate when language priors dominate weak or ambiguous visual evidence. Existing contrastive decoding methods mitigate this problem by comparing predictions from the original image with those from externally perturbed visual inputs, but such references can introduce off-manifold artifacts and require costly extra forward passes. We propose SIRA, a training-free internal contrastive decoding framework that constructs a counterfactual reference inside the same LVLM by exploiting the staged information flow of multimodal transformers. Instead of removing visual information from the input, SIRA first lets image and text tokens interact through a shared prefix, forming an aligned multimodal state that preserves prompt interpretation, decoding history, positional structure, and early visual grounding. It then forks a counterfactual branch in later transformer layers, where attention to image-token positions is masked. This branch retains the shared multimodal context but lacks continued access to fine-grained visual evidence, yielding a language-prior-dominated internal reference for token-level contrast. During decoding, SIRA suppresses tokens that remain strong without late visual access and favors predictions whose advantage depends on the full visual pathway. Experiments on POPE, CHAIR, and AMBER with Qwen2.5-VL and LLaVA-v1.5 show that SIRA consistently reduces hallucinations while preserving descriptive coverage and incurring lower overhead than two-pass contrastive decoding. SIRA requires no training, external verifier, or perturbed input, and applies to open-weight LVLMs with white-box inference access.

Tian Qin, Junzhe Chen, Yuqing Shi, Tianshu Zhang, Qiang Ju, Lijie Wen• 2026

Related benchmarks

Task	Dataset	Result
Object Hallucination Evaluation	MS-COCO (POPE Adversarial)	Accuracy87.83	205
Object Hallucination Evaluation	MS-COCO POPE (Popular)	Accuracy89.7	158
Caption Hallucination Evaluation	CHAIR	--	122
Object Hallucination Evaluation	MS-COCO POPE Random	Accuracy91.13	121
Object Hallucination Evaluation	A-OKVQA POPE Popular	Accuracy89.67	91
Object Hallucination Evaluation	A-OKVQA POPE Random	Accuracy92.1	75
Object Hallucination Evaluation	POPE GQA Popular	Accuracy88.37	70
Object Hallucination Assessment	A-OKVQA POPE (Adversarial)	Accuracy0.8363	57
Object Hallucination Evaluation	POPE-GQA Adversarial	Accuracy84.27	34
Multi-modal Hallucination Evaluation	AMBER	CHAIR4.8	28

Showing 10 of 11 rows

Other info

Follow for update

@wizwand_team Discord