Preprint Open access
Do Not Train Away Uncertainty: Early Uncertainty Anchored Calibration
Deep neural networks, including large language models, have achieved remarkable performance across various tasks. However, they are prone to overconfidence during training or fine-tuning. In this work, we observe a consistent phenomenon across different models that the early model is better calibrated, while later trai …