Research questionHow can vision-language models correct unsafe generations token by token without disrupting safe reasoning?Always-on safety mechanisms can alter decoding even when a generation is already safe, potentially degrading the model’s general multimodal capabilities. The challenge is to identify unsafe generation states as they arise and intervene without disturbing safe trajectories.