Research questionHow can unified multimodal models reduce redundant inference computation across understanding and generation without sacrificing quality?Unified models serve understanding and generation, but the computational contribution of tokens, layers, and timesteps varies across tasks. Removing this redundant work can degrade quality when the two tasks require different computation patterns.