CoVLM-Bench: A Real-World Benchmark for Cooperative Driving Question Answering and Planning
Vision-language models (VLMs) have made substantial progress in autonomous driving, but their success has primarily been studied in ego-centric scenes. Infrastructure-side observations provide views beyond the ego vehicle's field of view, yet conventional cooperative-driving systems typically transform them into geomet...