Skip to content
arXiv cs.CL · Papers

Do VLMs Align Better with Humans than LLMs during Natural Reading?

arXiv:2605.28818v2 Announce Type: replace Abstract: Large language models have become increasingly useful computational models of human language processing, but it remains open whether vision-language learning makes text representations more human-like during natural reading. We address this question by comparing match