Aug 2026· Communications in Transportation Research· 0 citations
TL;DR
Which2comm, a novel multi-agent 3D object detection framework leveraging object-level sparse features, consistently outperformed other state-of-the-art methods on both detection performance and communication cost, exhibiting superior robustness to real-world latency.
Abstract
Collaborative perception allows real-time inter-agent information exchange and thus offers invaluable opportunities to enhance the perception capabilities of individual agents. However, limited communication bandwidth in practical scenarios restricts the inter-agent data transmission volume. This implies a trade-off between perception performance and communication cost. To address this issue, we propose Which2comm, a novel multi-agent 3D object detection framework leveraging object-level sparse features. By integrating semantic information of objects into detection boxes, we introduce semantic detection boxes (SemDBs). Innovatively transmitting these object-level sparse features among agents not only significantly reduces the demanding communication volume, but also improves object detection performance. Moreover, an adaptive strategy is further proposed to select only safety-critical connected and automated vehicles (CAVs) for collaborative perception when there are multiple CAVs available, thereby maintaining stable communication costs. To validate the proposed method, a large-scale, multi-modal dataset, Multi-V2X, is established for vehicle-to-everything (V2X) perception tasks with various CAV penetration rates. Multi-V2X comprises 146k frames with over 4.2 million 3D annotations, featuring high agent density (up to 31 agents) to evaluate perception robustness in complex traffic environments. Extensive experiments demonstrate that Which2comm consistently outperformed other state-of-the-art methods on both detection performance and communication cost, exhibiting superior robustness to real-world latency.
Vehicle-to-everything (V2X) techniques expand the capability boundaries of connected and autonomous vehicles (CAVs). However, deploying V2X-enabled end-to-end autonomous driving (E2E-AD) systems still faces a trade-off between sharing high-resolution perception features and V2X communication bandwidth constraints. Furt...
Han Jiang, Zi-Yang Yan, Zechang Ye et al.· IEEE Transactions on Cogniti...· 0 citations
Cooperative perception through vehicle-to-vehicle (V2V) communication can resolve occlusions that single-vehicle systems cannot overcome, yet existing frameworks rely on simulated environments and expensive multi-sensor platforms. This paper presents CoSafe, a cooperative perception and reasoning framework built on rea...
Iosif-Alin Beti, P. Herghelegiu, C. Căruntu· Italian National Conference...· 0 citations
Cooperative perception can improve the situational awareness of connected and autonomous vehicles by exchanging complementary sensing information among nearby agents. However, systematic multi-vehicle cooperation also increases vehicle-to-everything (V2X) communication demand, processing overhead, and operational resou...
M. Alrashidi, Wajih Abdallah· Italian National Conference...· 0 citations
This work presents a conceptual framework for Collaborative Joint Perception and Prediction (Co-P&P) that improves motion prediction of surrounding road users, thereby enhancing situational awareness in complex and dynamic traffic environments.
Lei Wan, Hannan Ejaz Keen, Alexey V. Vinel· 0 citations
Cooperative perception enables connected autonomous vehicles to extend their sensing range and overcome occlusions by exchanging sensor data. However, its real-world deployment is hindered by asynchronous sensor streams and inaccurate localization of occluded regions. This work presents RAO++, a real-time cooperative p...
Ruiyang Zhu, Qingzhao Zhang, Xumiao Zhang et al.· ACM transactions on sensor n...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.