نسخة أولية وصول مفتوح
PermVLA: Factorization Order as a Regularizer for VLA Learning
Vision-language-action (VLA) policies commonly learn action chunks through a fixed left-to-right (LTR) factorization, although the same expert trajectory distribution admits many valid chain-rule factorizations. We identify factorization order as an overlooked regularization choice and introduce causally anchored permu …