نسخة أولية وصول مفتوح
Shared Experience, Separate Learning: Companion Confidence Calibration for LLMs
Reliable self-assessment is essential for large language models (LLMs), yet they often remain highly confident when their answers are wrong. We study \emph{concurrent confidence calibration}, where confidence is learned alongside capability improvement rather than calibrated only after training. Reinforcement learning …