الباحثون

Yufei Han

المنشورات 2

نسخة أولية وصول مفتوح

Blocking at the Boundary: Auditing Long-Horizon Agents against Staged Prompt Injection

Jingkai Liu, Yufei Han, Xiaoting Lyu وآخرون · 2026

Long-horizon agents consume external content, invoke tools, and modify persistent state. Indirect prompt injection can exploit task-specific context, propagate across causally connected stages, and alter a consequential action while the workflow continues; we term this staged prompt injection. We build an automated, fe …

نسخة أولية وصول مفتوح

SecProbe: Adaptive Evaluation of Coding Agents on Cybersecurity Vulnerabilities

Xiaonan Luo, Yue Huang, Kehan Guo وآخرون · 2026

Assessing cybersecurity vulnerability awareness in coding agents requires evaluations that reveal capability gaps and remain informative as models evolve. Static benchmarks offer fixed coverage and difficulty, while scarce vulnerable repositories and costly expert authoring limit their renewal at scale. We introduce Se …

المؤلفون المشاركون