Research questionHow can LLMs interleave reasoning with reliable step-level self-critique without a separate verifier?LLMs may produce plausible reasoning without detecting mistakes until the final answer. Separate verifiers can provide feedback, but they add coordination and system complexity, while self-critique must remain aligned with actual reasoning correctness.