WorldReward: Reward Modeling for Camera-Conditioned World Models
This work presents WorldReward, a VLM-based pairwise preference reward model that unifies action-consistency and visual-quality evaluation for camera-conditioned world models, and introduces WorldReward-Bench, a human-annotated benchmark measuring reward-model agreement with human preferences across action consistency,...