Measuring the Value of World-Model Updates: A Counterfactual Utility Protocol for Continual Adaptation
Continual world models must decide whether new data justify changing the model. Fixed replay schedules and prediction-error triggers specify when to update, but neither reveals the value of an individual update: one deployment run cannot show how the same model would have performed at that moment had it held its parame...