نسخة أولية وصول مفتوح
VersaCamVLA: Camera-Configurable VLA Policies for Robotic Manipulation
Vision-Language-Action (VLA) models have emerged as powerful foundations for robotic manipulation, but their reliance on fixed camera configurations during training makes them brittle to changes in camera count or pose during deployment. To overcome these limitations, we propose VersaCamVLA, a camera-configurable frame …