AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (1.4 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Research Article | Open Access | Just Accepted

Vehicle-Infrastructure Cooperative General Object Detection through Feature Flow and Differentiable Pose-based Spatial Alignment

Yanding Yang1,2,Rujun Yan1,Jinyu Miao1Zhiwei Zhao2Kewei Wang3Kun Jiang1( )Diange Yang1( )

1 School of Vehicle and Mobility, State Key Laboratory of Intelligent Green Vehicle and Mobility, Tsinghua University, Beijing 100084, China.

2 Dongfeng Motor Corporation Research & Development Institute, Wuhan 430100, China.

3 Dongfeng USharing Technology Co., Ltd, Wuhan 430058, China.

Yanding Yan and Rujun Yan contributed equally to this work.

Show Author Information

Abstract

Vehicle-infrastructure cooperative perception is critical for advanced autonomous driving but faces challenges: heavy reliance on ground-truth (GT) vehicle-infrastructure poses for spatial synchronization and limitations of traditional object detection in dynamic scene modeling. This paper proposes a differentiable pose-based spatial alignment (DPSA) module and explores occupancy flow output for cooperative perception. The DPSA module eliminates GT pose requirements by estimating 3-degree-of-freedom relative poses via feature fusion, spatial average pooling, and fully connected layers, balancing accuracy and practicality. Evaluations on DAIR-V2X and V2X-Seq datasets show DPSA outperforms misaligned fusion methods in key metrics, maintains robustness under low-location-precision scenarios, and reduces inference time. Notably, occupancy flow demonstrates superior dynamic modeling capability compared to traditional object detection by capturing spatial occupancy changes and motion flows (instead of bounding boxes), enabling high-accuracy future occupancy prediction and occlusion resistance via multi-source temporal context. This work advances practical cooperative perception through an efficient synchronization solution and highlights occupancy flow advantages, bridging algorithmic innovation and engineering deployment.

References

【1】
【1】
 
 
Journal of Intelligent and Connected Vehicles

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Yang Y, Yan R, Miao J, et al. Vehicle-Infrastructure Cooperative General Object Detection through Feature Flow and Differentiable Pose-based Spatial Alignment. Journal of Intelligent and Connected Vehicles, 2026, https://doi.org/10.26599/JICV.2026.9210091

344

Views

39

Downloads

0

Crossref

0

Web of Science

0

Scopus

Received: 07 January 2026
Revised: 07 June 2026
Accepted: 24 July 2026
Available online: 28 July 2026

© The Author(s) 2026.

This is an open access article under the terms of the Creative Commons Attribution 4.0 International License (CC BY 4.0,
http://creativecommons.org/licenses/by/4.0/).