Research questionHow can idle inference resources reduce scarce-GPU training cost without biasing gradient estimates?Training can face scarce, expensive GPU forwards even when lower-cost inference capacity is idle. Approximate gradients could reduce that bottleneck, but their errors may distort optimization if they introduce bias.