One email each morning — yesterday's AI, sortedGet it in your inbox

METAL LAB

Self- and Other-Labels Induce Bidirectional Bias in LLM Judges

arXiv:2608.180912026-08-20

As LLM-as-a-judge systems become increasingly widespread, self-preference in LLMs -- the tendency to favor one's own outputs -- raises growing concerns about evaluation reliability. However, it has been studied predominantly on generated text, where stylistic features and response quality are inevitably conflated. As a result, existing measurements cannot separate genuine self-preference from these confounds. We address this by changing the object

Authors · Songeun Chae, Min Kim, Donghoon Jung, Seojin Choi, Seohyon Jung

Read on arXiv

Latest papers

All papers →

Latest from METAL LAB