الباحثون

Hanyu Wang

المنشورات 3

نسخة أولية وصول مفتوح

Balancing Reference Guidance and Free Generation in Trajectory Rollouts for Reasoning RL

A verified reference solution provides a correct trajectory for training a reasoning model. Alternatively, a prefix of the reference can guide the model in generating a trajectory of its own. How much reference guidance should we provide? We study this question through prefix continuation, where the model continues fro …

نسخة أولية وصول مفتوح

OverAct: Measuring and Mitigating Proactive Over-Authorization in LLM Tool-Calling Agents

Taolin Zhang, Jiuheng Wan, Hanyu Wang وآخرون · 2026

LLM agents with tool-calling capabilities can access external services and private user data, but they may retrieve more information than a user's request explicitly requires. We study this behavior in structured tool-calling agents and term it proactive over-authorization. This setting differs from filesystem-level co …

نسخة أولية وصول مفتوح

How the Audit Rule Shapes Faithful Factor Explanations in LLMs

Taolin Zhang, Hanyu Wang, Jiuheng Wan وآخرون · 2026

Large language models are often asked which input factors influenced their outputs. For structured inputs, such reports can be checked by counterfactual perturbation, but each factor must be queried multiple times to estimate its effect, so verification is usually budget-limited. We study how this limited-budget settin …

المؤلفون المشاركون