arXiv cs.CL
· Papers
Measuring the Cross-Lingual Comprehension Gap: How the language of the evidence shapes what language models understand
arXiv:2608.06506v1 Announce Type: new Abstract: Language models are often evaluated as though capabilities demonstrated in English remain equally available when the same content is presented in other languages. Traditional multilingual benchmarks rarely isolate language while holding content, question, reference answer