Research questionHow can active preference learning obtain scalable, calibrated uncertainty for neural reward models without full Bayesian inference?Active preference learning must choose which comparisons to request, but reliable uncertainty estimates become expensive for neural reward models when inference considers all parameters. Poorly calibrated uncertainty can lead to less informative queries and inefficient reward learning.