매일 아침, 어제의 AI를 한 통으로 정리해 보내드립니다메일로 받아보기

METAL LAB

Why Do AI Agents Break Rules? How Framing, Context, and Social Signals Shape Compliance

arXiv:2608.123232026-08-14

arXiv:2608.12323v1 Announce Type: new Abstract: Specifying a penalty can paradoxically convert a legal obligation into a cost-benefit calculation that favors violation. We demonstrate that this enforcement information paradox systematically occurs in AI agents. While most AI safety evaluations test whether models fail, we investigate why, applying compliance theory from law and economics as a diagnostic tool. We treat compliance theories not as metaphors but as empirical hypotheses and show that

저자 · Mika Okamoto, Ansel Kaplan Erol, Kutluhan Erol

arXiv에서 원문 보기