Abstract

Latent world models offer a promising way to improve Vision-Language-Action policies by capturing the consequences of actions. However, models trained primarily on expert demonstrations have limited exposure to failure outcomes and may struggle to distinguish visually similar successful and failed interactions. We propose \textbf{WorldGuide}, a framework that learns these distinctions in latent space and uses them to guide policy training. WorldGuide combines predictive pretraining on successful and failed trajectories with contrastive learning on matched success--failure pairs. The learned predictor then provides a differentiable reward to guide joint optimization of the policy and visual encoder. The predictor is discarded after training, so deployment requires no additional world-model inference. Extensive experiments show that WorldGuide substantially improves VLA reliability and achieves state of the art performance on LIBERO 100 and SimplerEnv, reaching \textbf{96.8\%} and \textbf{72.0\%}, respectively. Code will be publicly available.

Keywords

Subject

Publication details

Journal
Not available
Open access
Green open access

Cite this article

APA 7

Liu, L., Zhang, L., Song, Z., Yang, W., Zhuang, Y., Zhuge, Y., Tao, S., Liu, W., & Lu, H. (2026). WorldGuide: Learning Success-Failure Boundaries in Latent World Models for Vision-Language-Action Policies. https://omanscience.com/en/articles/worldguide-learning-success-failure-boundaries-in-latent-world-models-for-vision-language-action-policies

MLA 9

Liu, Lin, et al. "WorldGuide: Learning Success-Failure Boundaries in Latent World Models for Vision-Language-Action Policies." https://omanscience.com/en/articles/worldguide-learning-success-failure-boundaries-in-latent-world-models-for-vision-language-action-policies.

Chicago (author–date)

Liu, Lin, Lu Zhang, Ziying Song, Wu Yang, Yuzheng Zhuang, Yunzhi Zhuge, Shuai Tao, Wulong Liu, and Huchuan Lu. 2026. "WorldGuide: Learning Success-Failure Boundaries in Latent World Models for Vision-Language-Action Policies." https://omanscience.com/en/articles/worldguide-learning-success-failure-boundaries-in-latent-world-models-for-vision-language-action-policies.

Harvard

Liu, L., Zhang, L., Song, Z., Yang, W., Zhuang, Y., Zhuge, Y., Tao, S., Liu, W. and Lu, H. (2026) 'WorldGuide: Learning Success-Failure Boundaries in Latent World Models for Vision-Language-Action Policies', Available at: https://omanscience.com/en/articles/worldguide-learning-success-failure-boundaries-in-latent-world-models-for-vision-language-action-policies.

Vancouver

Liu L, Zhang L, Song Z, Yang W, Zhuang Y, Zhuge Y, et al. WorldGuide: Learning Success-Failure Boundaries in Latent World Models for Vision-Language-Action Policies. https://omanscience.com/en/articles/worldguide-learning-success-failure-boundaries-in-latent-world-models-for-vision-language-action-policies

IEEE

L. Liu, L. Zhang, Z. Song, W. Yang, Y. Zhuang, Y. Zhuge, S. Tao, W. Liu, and H. Lu, "WorldGuide: Learning Success-Failure Boundaries in Latent World Models for Vision-Language-Action Policies," https://omanscience.com/en/articles/worldguide-learning-success-failure-boundaries-in-latent-world-models-for-vision-language-action-policies.