arXiv cs.LG
· Papers
CaliDist: Calibrating Large Language Models via Behavioral Robustness to Distraction
arXiv:2606.05799v2 Announce Type: replace Abstract: Existing calibration methods for Large Language Models (LLMs) often overlook a critical dimension of trustworthiness: a model's behavioral robustness to irrelevant or misleading information. In this paper, we argue that a model's true confidence should reflect its sta