Research questionHow can speech deepfake detectors focus on synthesis artifacts while generalizing to unseen speakers?Speech encoder representations can carry strong speaker information, allowing detectors to learn speaker-specific correlations instead of cues from speech synthesis. This dependence can make detection unreliable when a speaker was not represented during training.