Revisiting Adversarial Robustness in Large-Scale 3-D Vision–Language Models
Focusing on zero-shot classification, this study demonstrates that 3D vision-language models exhibit heightened sensitivity to small coordinate perturbations, highlighting the need for a more rigorous security evaluation of 3D vision-language models.