Skip to content
arXiv cs.AI · Papers

Epistemic-aware Vision-Language Foundation Model for Fetal Ultrasound Interpretation

arXiv:2510.12953v4 Announce Type: replace-cross Abstract: Recent medical vision-language models have shown promise on tasks such as VQA, report generation, and anomaly detection. However, most are adapted to structured adult imaging and underperform in fetal ultrasound, which poses challenges of multi-view image reason