Research questionCan chain-of-thought monitoring detect preferences received through tools or inferred from raw artifacts?Chain-of-thought monitoring assumes that a model’s reasoning trace reveals the information influencing its answer. Preferences delivered through tool returns or inferred from unprocessed artifacts may affect answers without being clearly verbalized in the trace.