الباحثون

Chenlong Yin

المنشورات 1

نسخة أولية وصول مفتوح

Climbing the Hill: Prompt Injection Red-Teaming Against Frontier Models with Curriculum Reinforcement Learning

Chenlong Yin, Xiaolong Jin, Wei Zou وآخرون · 2026

Prompt injection is a leading security risk for LLMs and LLM-based applications such as agents. State-of-the-art red-teaming methods for prompt injection leverage reinforcement learning (RL) to train an attacker LLM to generate effective injected prompts. However, when targeting frontier LLMs such as GPT-6-Luna, a majo …

المؤلفون المشاركون