AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (8.8 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Research Article | Open Access

Which2comm: Efficient collaborative perception framework with connected and automated vehicles

Duanrui Yu1,D, Anqi Qu1,D, Jing You1,D, Dingyu Wang1, Rongsong Li1, Shaocheng Jia2( ), Xin Pei1( )
Department of Automation, BNRist, Tsinghua University, Beijing 100084, China
Institute of Intelligent Transportation Systems, College of Civil Engineering and Architecture, Zhejiang University, Hangzhou 310058, China

Duanrui Yu, Anqi Qu, and Jing You contributed equally to this work.

Show Author Information

Abstract

Collaborative perception allowed real-time interagent information exchange and thus offered invaluable opportunities to enhance the perception capabilities of individual agents. However, limited communication bandwidth in practical scenarios restricted the interagent data transmission volume. This implied a trade-off between perception performance and communication cost. To address this issue, we proposed Which2comm, a novel multiagent three-dimensional (3D) object detection framework leveraging object-level sparse features. By integrating semantic information of objects into detection boxes, we introduced semantic detection boxes (SemDBs). Innovatively transmitting these object-level sparse features among agents not only significantly reduces the demanding communication volume but also improves object detection performance. Moreover, an adaptive strategy was further proposed to select only safety critical connected and automated vehicles (CAVs) for collaborative perception when there were multiple CAVs available, thereby maintaining stable communication costs. To validate the proposed method, a large-scale, multimodal dataset, Multi-V2X, is established for vehicle-to-everything perception tasks with various CAV penetration rates. Multi-V2X comprises 1.46 ×105 frames with over 4.2 × 106 3D annotations, featuring high agent density (up to 31 agents) to evaluate perception robustness in complex traffic environments. Extensive experiments demonstrate that Which2comm consistently outperforms other state-of-the-art methods on both detection performance and communication cost, exhibiting superior robustness to real-world latency.

Graphical Abstract

References

【1】
【1】
 
 
Communications in Transportation Research
Article number: 9640048

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Yu D, Qu A, You J, et al. Which2comm: Efficient collaborative perception framework with connected and automated vehicles. Communications in Transportation Research, 2026, 6(3): 9640048. https://doi.org/10.26599/COMMTR.2026.9640048

397

Views

44

Downloads

0

Crossref

0

Web of Science

0

Scopus

Received: 06 March 2026
Revised: 20 June 2026
Accepted: 17 August 2026
Published: 30 September 2026
© The Author(s) 2026.

This is an open access article under the terms of the Creative Commons Attribution 4.0 International License (CC BY 4.0 http://creativecommons.org/licenses/by/4.0/).