الباحثون

Hengyi Zhu

المنشورات 1

نسخة أولية وصول مفتوح

Imagine the Future, Internalize the Gist: Efficient VLA Reasoning via Internalized Spatiotemporal Imagination

Shenglan Li, Zhendong Mi, Hengyi Zhu وآخرون · 2026

Vision-language-action (VLA) models increasingly incorporate intermediate reasoning to improve robotic manipulation, yet existing approaches primarily reason about observed states without explicitly anticipating future scene evolution. Extending such reasoning to explicit future rollouts at every inference step, howeve …

المؤلفون المشاركون