Research questionHow can vision-language-action policies follow execution details beyond a robot task’s goal?Robot trajectories are often labeled only with coarse task goals, leaving choices such as the active arm, approach direction, and contact region unspecified. Without language grounding for these choices, policies are difficult to steer during otherwise successful task execution.