الباحثون

Pingzhi Li

المنشورات 2

نسخة أولية وصول مفتوح

Test-Time Unlearning via Sparse Autoencoder

Pingzhi Li, Jinhao Duan, Vaishnav Tadiparthi وآخرون · 2026

Machine unlearning aims to remove specific knowledge from a trained large language model (LLM) without retraining from scratch. Existing methods modify model weights via gradient ascent and its advances. While effective on certain benchmarks, these weight-based approaches exhibit a sharp forget-utility trade-off, where …

المؤلفون المشاركون