نسخة أولية وصول مفتوح
Train Together or Merge Later? Unifying VLA Experts via a Shared Action Interface
Co-training offers a straightforward way to build a multi-task vision-language-action (VLA) policy, but can fall short of the performance achieved by training each task independently. The challenge is to retain these task-specific gains in a multi-task policy without joint post-training. Combining independently trained …