Research questionHow can vision-language-action policies act reliably when irrelevant sensors are corrupted or only one informative sensor remains?Limited, homogeneous robot demonstrations can make a VLA policy depend on cross-sensor correlations rather than the sensor carrying task-relevant information. Consequently, corrupting an irrelevant modality can disrupt actions, while the policy may fail when only one informative modality remains.