نسخة أولية وصول مفتوح
Residual Modeling Closes the Regression and Generative Policy Gap in Robot Learning
Learning from demonstration has enabled impressive robot behaviors. A common choice for policy learning is to use diffusion or flow matching (Flow-Policies), which often outperforms direct action regression trained with mean squared error (MSE-Policies). This gap is commonly attributed to multimodal demonstrations. We …