Research questionHow can discrete audio tokenizers preserve semantics and acoustic fidelity for both understanding and generation?Audio tokenizers must compress continuous sound into manageable discrete sequences without losing either high-level content or acoustic detail. Separate semantic and acoustic streams can also introduce redundancy or misalignment when the same representation must support analysis and synthesis.