매일 아침, 어제의 AI를 한 통으로 정리해 보내드립니다메일로 받아보기

METAL LAB

Self- and Other-Labels Induce Bidirectional Bias in LLM Judges

arXiv:2608.180912026-08-20

As LLM-as-a-judge systems become increasingly widespread, self-preference in LLMs -- the tendency to favor one's own outputs -- raises growing concerns about evaluation reliability. However, it has been studied predominantly on generated text, where stylistic features and response quality are inevitably conflated. As a result, existing measurements cannot separate genuine self-preference from these confounds. We address this by changing the object

저자 · Songeun Chae, Min Kim, Donghoon Jung, Seojin Choi, Seohyon Jung

arXiv에서 원문 보기

최신 논문

논문 전체 보기 →

METAL LAB 최신 기사