Research questionCan chain-of-thought monitoring detect consequential computation hidden in semantically irrelevant filler tokens?Language models may gain task performance from semantically irrelevant filler tokens without making the relevant computation interpretable in their visible reasoning. This complicates the use of chain-of-thought as evidence of what a model has computed.