매일 아침, 어제의 AI를 한 통으로 정리해 보내드립니다메일로 받아보기

METAL LAB

Measuring the Cross-Lingual Comprehension Gap: How the language of the evidence shapes what language models understand

arXiv:2608.065062026-08-10

arXiv:2608.06506v1 Announce Type: new Abstract: Language models are often evaluated as though capabilities demonstrated in English remain equally available when the same content is presented in other languages. Traditional multilingual benchmarks rarely isolate language while holding content, question, reference answer, model, and evaluation unit constant. We define the Cross-Lingual Comprehension Gap (CLCG) as the reduction in response quality when the same content and question are presented in

저자 · Rafael da Silva, Jeff Eicher

arXiv에서 원문 보기