الباحثون

Kezhen Chen

المنشورات 4

نسخة أولية وصول مفتوح

Suppressing Pressure, Amplifying Evidence: Self-Guided Attention Steering to Mitigate Sycophancy and Stubbornness

Yinghao He, Mengyu Xu, Haixiang Sun وآخرون · 2026

Reliable language models should resist unsupported user pressure while effectively using objective contextual information. However, models may exhibit sycophancy by yielding to unsupported user pressure or contextual stubbornness by failing to update their answers when relevant contextual information warrants revision. …

نسخة أولية وصول مفتوح

On Unlearning for Time-series Forecasting

Zeyu Shi, Yanhui Luo, Ziming Hong وآخرون · 2026

Time-series forecasting is widely used in sensitive domains. Models in these settings are often trained on longitudinal user- or entity-level records, which may later require removal because they contain sensitive or proprietary information or have been corrupted by sensor failures. To address such deletion requests wi …

نسخة أولية وصول مفتوح

Storage Is Not Strategy: State-Conditioned Support Control for LLM Unlearning

Tianhao Qian, Ziming Hong, Chongyang Gao وآخرون · 2026

Many localized large language model (LLM) unlearning methods select a small parameter subset from a localization signal and keep it fixed during optimization. The parameters most associated with a target, however, need not be the best ones to update, and candidate interventions can change value as optimization proceeds …

المؤلفون المشاركون