الباحثون

Xinyuan Qian

المنشورات 2

نسخة أولية وصول مفتوح

FragToken: Amplifying LLM Inference Costs through Noncanonical Token Generation

Zihan Wang, Rui Zhang, Xinyuan Qian وآخرون · 2026

As large language model (LLM) inference becomes increasingly expensive, resource-consumption attacks pose a growing threat to model providers. Existing attacks typically amplify cost by inducing abnormally long or repetitive outputs on attacker-controlled or triggered requests, making them easier to detect and limiting …

المؤلفون المشاركون