Research questionHow can we detect object hallucinations in vision-language models when visual grounding shifts across layers?A model may mention objects that are absent from an image, and its visual grounding can change substantially across internal layers. Reliable detection therefore requires distinguishing predictions supported by visual evidence from hallucinated ones.