الباحثون

Zhaowei Liang

المنشورات 1

نسخة أولية وصول مفتوح

ReF-HIL: Shaping the Critic around Human Action Neighborhoods for Efficient Human-in-the-Loop Reinforcement Learning

Shaoyin Luo, Song Wang, Shibo Xia وآخرون · 2026

Human-in-the-loop reinforcement learning (HIL-RL) offers a promising route to efficient training of robotic manipulation policies by combining autonomous learning with human demonstrations and online corrections. However, insufficient use of successful human experience in value learning prolongs costly real-world train …

المؤلفون المشاركون