Research questionDo refusal representations transfer across language-model architectures, and where should safety interventions read them?Architectures differ in how they mix and update token information, so a refusal signal identified in one model may not be directly usable in another. Safety tooling must determine both whether the representation transfers and where the relevant computation is exposed.