每天早上一封邮件,把昨天的 AI 梳理好订阅邮件

METAL LAB

Self- and Other-Labels Induce Bidirectional Bias in LLM Judges

arXiv:2608.180912026-08-20

As LLM-as-a-judge systems become increasingly widespread, self-preference in LLMs -- the tendency to favor one's own outputs -- raises growing concerns about evaluation reliability. However, it has been studied predominantly on generated text, where stylistic features and response quality are inevitably conflated. As a result, existing measurements cannot separate genuine self-preference from these confounds. We address this by changing the object

作者 · Songeun Chae, Min Kim, Donghoon Jung, Seojin Choi, Seohyon Jung

在 arXiv 阅读

最新论文

全部论文 →

METAL LAB 最新报道