Skip to content

Author

Cordelia Schmid

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Sep 2026

What do VLM-Based Vision-Language Navigation Models Rely on: Interpreting and Steering Policy Behavior

This work uses intervention-based metrics that measure how visual observations, instructions, and visual memory causally influence navigation decisions and shows that these navigation policies are sensitive to all input modalities and do not depend on a single one.

Débora Oliveira Makowski, Samiran Gode, Abhijeet Nayak et al. · 0 citations
Preprint Sep 2026

Spatially Aware World Action Model via Geometric Latent Diffusion

A Spatially Aware World Action Model (SA-WAM), which repurposes a pretrained video model for joint action, RGB, and depth prediction, enabling 3D-aware world modeling and action prediction within a single diffusion backbone.

Javier Alejandro Lopetegui Gonzalez, Paul Pacaud, Cordelia Schmid · 2 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.