Preprint Open access
An Empirical Study and Open Testbed for Federated Fine-Tuning of Vision-Language-Action Models
Adapting a pretrained Vision-Language-Action (VLA) model to a new robot, environment, or task requires demonstrations that are collected locally and often discarded. Federated learning is a promising approach to exploiting such distributed demonstrations by learning a shared policy. However, whether it can adapt large …