Preprint Open access
Alleviating Hallucination in Reasoning Tasks with Training-Free Uncertainty-Guided Steering
Recent work on hallucination detection in large language models has shown that, for a fixed pre-trained model and reasoning task, it is possible to estimate the model's confidence in the correctness of its outputs. Such uncertainty estimates have primarily been used to improve truthfulness by detecting or filtering con …