Research questionHow can model distillation block hidden teacher-trait transfer through clean data without degrading the target task?A teacher’s hidden bias can create subtle preference gaps in otherwise ordinary training outputs. Supervised fine-tuning may accumulate updates from those gaps, causing the student to adopt the teacher’s behavior even when the data reveals no obvious trait.