الباحثون

Luke Simon

المنشورات 2

نسخة أولية وصول مفتوح

Iterative Policy Refinement through Semantic Rollout Analysis

Feiyu Gavin Zhu, Qi Xu, Zhifei Deng وآخرون · 2026

Structured policies improve efficiency, robustness, and interpretability in imitation learning by introducing task-specific inductive bias, but existing structure generation methods rely either on extensive human input or on static domain knowledge encoded in LLMs, which may be inconsistent with the expert demonstratio …

المؤلفون المشاركون