매일 아침, 어제의 AI를 한 통으로 정리해 보내드립니다메일로 받아보기

METAL LAB

Every Token Counts: Exact Likert-Scale Distributions for Measuring LLM Attitudes and Biases

arXiv:2608.105032026-08-12

arXiv:2608.10503v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed as autonomous agents, accurately evaluating their latent values and biases is critical. The NLP community typically evaluates models using large, unstructured benchmarks. While effective for general capabilities, these datasets fundamentally conflate causal mechanisms: even when an aggregate bias is detected, unstructured evaluations cannot disentangle whether it stems from baseline traits,

저자 · Davood Wadi, Mohsen Ghodrat, Matthew Philp

arXiv에서 원문 보기