A benchmark should deliver more than a scalar score: what makes an evaluation trustworthy is the reasoning that justifies the score. This is especially critical for world models, where judging a rollout requires understanding whether physics, causality, and world state evolve correctly. Humans spot such violations natu...
Weiliang Chen, Haowen Sun, Jun Gao et al.· 2 citations
Improved text-visual attention patterns are introduced to enhance the fidelity of query-aware vision token selection and the Attention Gravity effect is correct, and a rank-based strategy to adaptively determine the sparsification ratio for each layer is introduced.
Yuan Zhang, Junpeng Ma, Qizhe Zhang et al.· IEEE Transactions on Pattern...· 2 citations
GaussianDet3D is presented, the first method to apply 3D Gaussian Splatting from multi-view images to 3D object detection in the context of autonomous driving, treating predicted Gaussian primitives as a pseudo-LiDAR point cloud fed into a sparse LiDAR detector.
Malaz Tamim, Wenzhao Zheng, Johannes Meier et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.