Preprint Open access
InfraVLA: Extending Vision-Language-Action Navigation with Infrastructure Cameras
Many indoor environments in which robots operate, such as warehouses, offices, and hospitals, already have cameras installed. They observe parts of the building that the robot cannot see from where it stands, yet navigation policies, including recent vision-language-action (VLA) models, do not use them. We propose Infr …