Targeting the Attention Heads Behind Object Hallucination in LLaVA
This story is from 2026-08-27. It is preserved in the archive; the latest stories are on the live feed.
arXiv:2608.24966v1 Announce Type: new Abstract: Vision-language models such as LLaVA-1.5-7B often hallucinate objects absent from the image when generating captions. We ask whether an interpretability diagnosis of this failure can guide a targeted fix, and we measure what that fix actually changes.…
Read the full story at arXiv cs.CV ↗
Timeline · 1 report
- 2026-08-27 04:00 · arXiv cs.CV
Targeting the Attention Heads Behind Object Hallucination in LLaVA