Abstract
Supervised fine-tuning (SFT) learns most aggressively from tokens that the model deems least likely. This helps acquire new behaviors, but also amplifies noisy or conflicting supervision and can overwrite useful pretrained knowledge. Through a unified policy-loss view, we revisit existing token-reweighting methods and show that they assign nonnegative coefficients to demonstrated tokens. Consequently, they can suppress or amplify supervised updates, but cannot reverse harmful features once learned. Moreover, larger training weights do not amount to feature extrapolation, since they change the optimization trajectory rather than scale a fixed SFT direction. We argue that reversal and extrapolation require a stable reference frame defined by a fixed SFT delta. Motivated by this, we propose SCALE (Selective Control of Adaptation via Local Entropy), an entropy-guided adaptation-strength-control method that freezes the pretrained model and the SFT delta and learns bounded token- and module-specific gates by minimizing predictive entropy alone. These gates suppress, reverse, or extrapolate frozen SFT features according to their alignment with entropy reduction. Across Qwen2.5-Math-1.5B, Qwen2.5-Math-7B, and Qwen3-4B-Base, SCALE achieves mathematical-reasoning averages of 37.84, 43.60, and 36.57, exceeding the strongest corresponding baselines while remaining competitive on general-retention benchmarks. It also attains the best average code-generation performance across HumanEval, HumanEval+, and MBPP for all three models. These results suggest that effective SFT correction can benefit from controlling how already learned residuals are used, rather than only modifying how they are learned.
Keywords
Subject
Publication details
- Journal
- Not available
- Open access
- Green open access
Cite this article
APA 7
Li, C., He, H., Gao, Y., Li, M., Ye, J., Yang, Q., & Ye, P. (2026). Rethinking Token Reweighting for SFT: Suppress, Reverse, and Extrapolate Learned Features. https://omanscience.com/en/articles/rethinking-token-reweighting-for-sft-suppress-reverse-and-extrapolate-learned-features
MLA 9
Li, Cunchun, et al. "Rethinking Token Reweighting for SFT: Suppress, Reverse, and Extrapolate Learned Features." https://omanscience.com/en/articles/rethinking-token-reweighting-for-sft-suppress-reverse-and-extrapolate-learned-features.
Chicago (author–date)
Li, Cunchun, Haonan He, Yifan Gao, Minglei Li, Jingqi Ye, Qingyu Yang, and Peng Ye. 2026. "Rethinking Token Reweighting for SFT: Suppress, Reverse, and Extrapolate Learned Features." https://omanscience.com/en/articles/rethinking-token-reweighting-for-sft-suppress-reverse-and-extrapolate-learned-features.
Harvard
Li, C., He, H., Gao, Y., Li, M., Ye, J., Yang, Q. and Ye, P. (2026) 'Rethinking Token Reweighting for SFT: Suppress, Reverse, and Extrapolate Learned Features', Available at: https://omanscience.com/en/articles/rethinking-token-reweighting-for-sft-suppress-reverse-and-extrapolate-learned-features.
Vancouver
Li C, He H, Gao Y, Li M, Ye J, Yang Q, et al. Rethinking Token Reweighting for SFT: Suppress, Reverse, and Extrapolate Learned Features. https://omanscience.com/en/articles/rethinking-token-reweighting-for-sft-suppress-reverse-and-extrapolate-learned-features
IEEE
C. Li, H. He, Y. Gao, M. Li, J. Ye, Q. Yang, and P. Ye, "Rethinking Token Reweighting for SFT: Suppress, Reverse, and Extrapolate Learned Features," https://omanscience.com/en/articles/rethinking-token-reweighting-for-sft-suppress-reverse-and-extrapolate-learned-features.