الباحثون

Yuetong Liu

المنشورات 1

نسخة أولية وصول مفتوح

Safe Actions Alone Do Not Ensure Safe Agents: Identifying Unfulfilled Obligations with Guard Models

Youwei Feng, Yitong Zhang, Yuetong Liu وآخرون · 2026

Guard models are increasingly used to safeguard LLM-based agents, primarily by identifying actions that agents are forbidden to perform. However, identifying forbidden actions alone is insufficient to ensure agent safety. In this paper, we argue that agent safety also depends on identifying required yet unperformed saf …

المؤلفون المشاركون