Research questionHow can high-resolution diffusion Transformers prune tokens without sacrificing image fidelity or predictable compute?At high resolution, self-attention cost grows quadratically with the number of image tokens. Pruning tokens can reduce this burden, but removing information may harm generated-image fidelity and make computation harder to predict.