Skip to content

Author

Jingkai Wang

3 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Sep 2026

Beyond Retrieval Relevance: Scene-Grounded Risk Entailment for Vision-Language Driving

Retrieval-augmented generation (RAG) gives vision--language driving systems access to external safety knowledge, yet a retrieved risk rule may be relevant without applying to the current scene. A vision--language model (VLM) receiving such knowledge must ground objects, bind entities across time, and verify relations b...

Jia-Xin Liu, Rui-Lin Yu, Liang Peng et al. · 0 citations
Preprint Aug 2026

Beyond Instance Slots: Semantically Rich World Models for Physical Interaction Planning

The Semantically Rich World Model is presented, a task-conditioned world model structured around five functional roles: gripper, target, goal, relation, and phase, which transforms object-centric prediction into a semantic interface linking visual dynamics with planning-oriented decision making.

Juntao Cheng, Jingkai Wang, Yi-Jun Shen et al. · 1 citation
Preprint Aug 2026

SLIM-0.5B: Learning Action-Grounded Predictive Latents for Robot Manipulation

SLIM (Self-supervised Latent Interaction Model), a compact 0.5B-parameter latent interaction policy, which matches or exceeds representative large-scale VLA and world-action-model baselines with fewer parameters, no additional embodied pretraining, lower inference latency, and substantially lower GPU memory usage.

Jing-Kai Wang, Zihan Tang, Gu Zhang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.