Real-world driving is inherently multi-agent, yet most existing driving world models generate observations from a single ego vehicle. Independently extending them to multiple vehicles does not ensure that different agents observe a consistent shared world. We present CoDrive, a cross-vehicle, multi-view driving video g...
Yu Meng, Bai-Ning Zhao, Jun-Tao Wu et al.· 0 citations
IMPACT is introduced, a scalable Interaction-aware Model training framework with Prior-guided Attention Calibration and Targeting, which consistently outperforms the corresponding MSE-trained baselines, improving interaction fidelity, physical plausibility, and visual quality.
Rong-Ze Tang, Jianjie Fang, Zhao-Lu Wang et al.· 1 citation
Causal Action Effect Reweighting (CAER), a general training paradigm that redistributes supervision toward the tokens whose predicted future is causally affected by the action, is introduced.
Jian-Jie Fang, Xvyuan Liu, Zi-You Wang et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.