Skip to content
arXiv cs.LG · Papers

CaliDist: Calibrating Large Language Models via Behavioral Robustness to Distraction

arXiv:2606.05799v2 Announce Type: replace Abstract: Existing calibration methods for Large Language Models (LLMs) often overlook a critical dimension of trustworthiness: a model's behavioral robustness to irrelevant or misleading information. In this paper, we argue that a model's true confidence should reflect its sta