Skip to content
arXiv cs.CV · Papers

Confident but Unreliable: A Behavioral Safety Audit of Vision-Language Models on Brain MRI

arXiv:2608.02790v1 Announce Type: new Abstract: Vision-language models (VLMs), including medical specialists, are increasingly proposed for medical imaging, yet their stated confidence is rarely evaluated separately from correctness. We use brain MRI as a controlled, high-stakes testbed for a broader failure mode in fr