الباحثون

Hanwei Wu

المنشورات 1

نسخة أولية وصول مفتوح

Sequential Functional Structured Tucker Compression for Large Language Model Attentions

Jiangfeng Chen, Xinyu Wang, Tianshuo Yan وآخرون · 2026

Post-training compression of LLM attention is often formulated as independent matrix approximation, ignoring both the shared structure among attention projections and the representation shift introduced by earlier compression. We propose FTC, a sequential structured compression framework that adapts the approximation t …

المؤلفون المشاركون