Abstract

Vehicle-infrastructure cooperation can complement onboard sensing with broader and more informative observations of the traffic environment, providing valuable support for end-to-end autonomous driving. However, existing cooperative driving methods mainly exploit roadside information to enhance the representation of the current scene, while the future consequences of prospective driving actions are rarely modeled explicitly. This limits the ability of the planner to anticipate how its decisions may interact with the evolving traffic environment. To address this issue, we propose V2X-WAM, a cooperative world action model that tightly couples cooperative scene understanding, action generation, and future-world reasoning. V2X-WAM constructs a reliability-aware spatiotemporal representation from vehicle- and infrastructure-side observations, while compressing infrastructure information into a compact quantized message for efficient communication. Based on the resulting cooperative representation, a multimodal planner generates prospective trajectories, which explicitly condition future occupancy and dynamic-flow prediction. The predicted world consequences are then fed back to refine the planned trajectory, forming a closed interaction between action and future-world evolution. Experiments on a large-scale real-world cooperative driving dataset demonstrate that V2X-WAM consistently improves planning accuracy and safety over representative end-to-end cooperative driving methods, while achieving stronger future-world prediction and substantially lower communication overhead. Ablation studies further validate the effectiveness of the proposed design.

Keywords

Publication details

Journal
Not available
Open access
Green open access

Cite this article

APA 7

You, J., Tang, W., Wang, C., Zhao, Y., Hua, J., Shi, H., Zhang, W., Wang, L., & Ran, B. (2026). V2X-WAM: A Cooperative World Action Model for End-to-End Autonomous Driving. https://omanscience.com/en/articles/v2x-wam-a-cooperative-world-action-model-for-end-to-end-autonomous-driving

MLA 9

You, Junwei, et al. "V2X-WAM: A Cooperative World Action Model for End-to-End Autonomous Driving." https://omanscience.com/en/articles/v2x-wam-a-cooperative-world-action-model-for-end-to-end-autonomous-driving.

Chicago (author–date)

You, Junwei, Weizhe Tang, Can Wang, Yan Zhao, Jun Hua, Haotian Shi, Wei Zhang, Lin Wang, and Bin Ran. 2026. "V2X-WAM: A Cooperative World Action Model for End-to-End Autonomous Driving." https://omanscience.com/en/articles/v2x-wam-a-cooperative-world-action-model-for-end-to-end-autonomous-driving.

Harvard

You, J., Tang, W., Wang, C., Zhao, Y., Hua, J., Shi, H., Zhang, W., Wang, L. and Ran, B. (2026) 'V2X-WAM: A Cooperative World Action Model for End-to-End Autonomous Driving', Available at: https://omanscience.com/en/articles/v2x-wam-a-cooperative-world-action-model-for-end-to-end-autonomous-driving.

Vancouver

You J, Tang W, Wang C, Zhao Y, Hua J, Shi H, et al. V2X-WAM: A Cooperative World Action Model for End-to-End Autonomous Driving. https://omanscience.com/en/articles/v2x-wam-a-cooperative-world-action-model-for-end-to-end-autonomous-driving

IEEE

J. You, W. Tang, C. Wang, Y. Zhao, J. Hua, H. Shi, W. Zhang, L. Wang, and B. Ran, "V2X-WAM: A Cooperative World Action Model for End-to-End Autonomous Driving," https://omanscience.com/en/articles/v2x-wam-a-cooperative-world-action-model-for-end-to-end-autonomous-driving.