الباحثون

Lizhi Zhang

المنشورات 1

نسخة أولية وصول مفتوح

SkillPoison: Progressive Skill Poisoning via Successful Experiences

Lizhi Zhang, Xin He, Dianxuan Fu وآخرون · 2026

Self-improving LLM agents increasingly distill successful experiences into persistent, reusable skills. Existing skill attack methods corrupt this learning pipeline by injecting malicious triggers, behaviors, or false facts into individual experiences or extracted skills. However, such attacks are easily detected, and …

المؤلفون المشاركون