Just Noticeable Difference Modeling for Token Compression in Vision-Language-Action Models
Action-JND is introduced, which extends JND modeling to embodied perception by defining noticeability through the language-conditioned action response of a vision-language-action (VLA) policy in closed-loop control, and develops a lightweight token-wise JND estimator in deep visual-feature space to predict the maximum...