Research questionHow do attention and recurrent state divide factual retrieval from output style in hybrid language models?Hybrid language models retain context through both an attention cache and a fixed-size recurrent state, but it is unclear whether those channels store the same information. This makes it difficult to predict which channel supports exact recall versus broader output behavior.