الباحثون

Hang Yan

المنشورات 3

نسخة أولية وصول مفتوح

WEFT: Scaling Tool-Use Post-Training for General-Purpose Agents

Bo Mao, Hang He, Linting Wang وآخرون · 2026

Recent efforts to scale tool-use post-training have largely centered on the synthesis of executable environments, which constitute only one component of a broader agentic interaction system comprising the environment, task, agent harness, and evaluator. Scaling environments in isolation, however, does not guarantee com …

نسخة أولية وصول مفتوح

Does Learning to Predict the World Help Agents Act? Auditing World-Model Post-Training

Xinyu Che, Hang Yan, Yanchen Liu وآخرون · 2026

Predicting how an environment will change before acting is a natural route to better decision making for agents. Recent post-training methods therefore require agents to predict the next observation and turn that prediction into a reward or a direct supervision signal, which is called world model. Existing next-observa …

نسخة أولية وصول مفتوح

GameLogicBench: Evaluating Coding Agents on Runtime Game Logic with Tick-Level State Assertions

Xinyu Che, Yunfei Ge, Shihao Li وآخرون · 2026

Coding agents can modify and test code across large software projects. Game development is a domain where agents must implement gameplay rules. A game can end in a valid state even after violating its rules during the run. Current game-development benchmarks replay fixed examples, score videos, or ask another model to …

المؤلفون المشاركون