Skip to content

Author

Cordelia Schmid

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Sep 2026

What do VLM-Based Vision-Language Navigation Models Rely on: Interpreting and Steering Policy Behavior

This work uses intervention-based metrics that measure how visual observations, instructions, and visual memory causally influence navigation decisions and shows that these navigation policies are sensitive to all input modalities and do not depend on a single one.

Débora Oliveira Makowski, Samiran Gode, Abhijeet Nayak et al. · 0 citations
Preprint Sep 2026

Spatially Aware World Action Model via Geometric Latent Diffusion

World Action Models (WAMs) leverage the capabilities of large-scale pretrained video diffusion models to jointly predict future observations and actions, inheriting rich visual and physical priors from internet-scale video. This has made them a promising paradigm for robot policy learning, yet the prevailing models ope...

Javier Alejandro Lopetegui Gonzalez, Paul Pacaud, Cordelia Schmid · 2 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.