Skip to content
arXiv cs.CL · Papers

When Discourse Pressures Conflict: Information Structure in Vision-Language Model Outputs

arXiv:2605.28346v2 Announce Type: replace Abstract: Vision-language models (VLMs) are increasingly evaluated for whether they identify the right visual content, but little is known about whether they express such content in a discourse-appropriate form. We address this research gap using information structure (IS), tes