Research questionWhen are SFT, ICL, and KL-regularized RL mathematically equivalent, and how does test-time reasoning differ?These paradigms are often treated as fundamentally different, even though several can implement related posterior projections. Their relationship changes depending on whether the signal comes from likelihoods, rewards, or advantages and whether learning occurs in weights or at inference time.