Preprint Open access
Hierarchical Reinforcement Learning with Stable Temporal Abstraction for Language Model Agents
Hierarchical reinforcement learning improves long-horizon control by organizing primitive actions around persistent subgoals and assigning credit at multiple temporal scales. Recent hierarchical language agents bring these benefits to interactive tasks by explicitly separating subgoal planning from action execution. We …