SEP 3, 2026 · PREPRINT
Subspace Inference Enables Efficient Active Reward Learning from Preferences
arXiv
This is an unrefereed arXiv preprint presenting a novel computational method for reward model training, not a clinical trial or peer-reviewed evidence synthesis; it reports algorithmic improvements but has not undergone peer review.
Study details
InterventionPreferenceEKF: sequential Bayesian filterin…
ComparatorOther Bayesian deep learning approaches for…