매일 아침, 어제의 AI를 한 통으로 정리해 보내드립니다메일로 받아보기

METAL LAB

Hidden Language Consistency Phenomena in Reasoning LLMs

arXiv:2608.084472026-08-11

arXiv:2608.08447v1 Announce Type: new Abstract: Multilingual reasoning models are commonly evaluated by whether they arrive at the correct answer, but not by whether they preserve the intended language while reasoning and responding. This omission conceals important multilingual behaviors that emerge as tasks become harder. In this paper, we study task difficulty, task accuracy, thinking-language consistency (TC), and answer-language consistency (AC) across reasoning models using PolyMath benchm

저자 · Muhammad Ali Shafique, Kelly Marchisio

arXiv에서 원문 보기