الباحثون

Yichen Wu

المنشورات 3

نسخة أولية وصول مفتوح

STEPQuant: When and Where Errors Matter in Delta-Rule Recurrent State Quantization

Bingchen Yao, Haobo Xu, Haokun Lin وآخرون · 2026

Linear attention replaces growing KV caches with fixed-size recurrent states, yet these persistent states can become a substantial memory bottleneck under concurrent serving. Directly quantizing recurrent states to low precision often leads to severe accuracy degradation, as quantization errors propagate through succes …

نسخة أولية وصول مفتوح

Learning an Anchored Prompt Space for Continual Adaptation of Large Language Models

Rongguang Ye, Zhan Zhuang, Yichen Wu وآخرون · 2026

Continually adapting large language models requires acquiring new knowledge while preserving previously learned capabilities. Jointly adapting model parameters and task-specific soft prompts offers a promising solution, but faces two key limitations: historical prompts may become less effective as the model evolves, wh …

نسخة أولية وصول مفتوح

Rewarding Reasoning, Not Answers: Fixing and Bounding Test-Time Reinforcement Learning on Medical QA

Kailong Fan, Anqi Pu, Yichen Wu وآخرون · 2026

Test-time reinforcement learning adapts a model on its own unlabeled test set using majority-vote pseudo-labels and has shown strong results in mathematics. We show that this recipe collapses on medical multiple-choice QA: accuracy stagnates while output diversity rapidly declines. Through a controlled experiment that …

المؤلفون المشاركون