Skip to content
arXiv cs.CL · Papers

Consistent but Miscalibrated: Evaluating LLM Limitations for Risk Communication in Natural Language

arXiv:2607.03882v2 Announce Type: replace Abstract: LLMs are increasingly deployed as post-hoc explainers of AI-generated outputs, yet it remains unclear whether they can reliably communicate probabilistic information in natural language. For this role to be viable, models must produce identical verbal descriptions for