Preprint
Jul 2026
RENEW: Towards Learning World Models and Repairing Model Exploitation from Preferences
This work proposes to repair exploitation directly using human preferences over imagined rollouts, leveraging the strong intuitive physics that allows humans to easily spot egregious dynamics hallucinations in pretrained world models.
L. M. Bhamidipaty, M. Kochenderfer, S. Ramamoorthy
· 0 citations