Jul 2026
WCog-VLA: A Dual-Level World-Cognitive Vision-Language-Action Model for End-to-End Autonomous Driving
WCog-VLA is proposed, a novel dual-level World-Cognitive VLA framework that successfully bridges semantic world forecasting with generative world evolution to achieve proactive autonomous driving.
Xue-Run Yan, Zhe-Xi Lian, Nuoheng Zhang et al.
· arXiv.org · 1 citation