Jetson-PI is proposed, a method for efficient VLA deployment on onboard devices via Foresight-Aligned Asynchronous Correction that trains a lightweight future correction module that predicts future environment representation conditioned on committed actions, enabling the action expert to directly predict actions from t...
Zebin Yang, Qi Wang, Yun-He Wang et al.· arXiv.org· 2 citations
Vision-Language-Action (VLA) models have achieved impressive performance on diverse embodied tasks. However, deploying VLA models on low-power onboard devices, such as the Jetson Orin, remains challenging due to their high computational complexity, which leads to substantial inference latency and low control frequency....
Zebin Yang, Qi Wang, Yun-He Wang et al.· 1 citation
Mimir is introduced, a neuro-symbolic memory that separates world memory from task memory and dynamically grounds them before each action, substantially outperforming current closed-source models.
Haoming Xu, Zhen-Lin He, Heng-Yi Wang et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.