每天早上一封邮件,把昨天的 AI 梳理好订阅邮件

METAL LAB

Demystifying Agent Skills: Why They Work-Until They Don't

arXiv:2608.140362026-08-13

Skills have emerged as a practical and effective approach for enhancing LLM agents at inference time through structured packages of knowledge. However, existing evaluations largely measure whether skills improve aggregated task success, leaving a more fundamental question underexplored: \textbf{When do skills help, why do they work, and where do they fail?} Through controlled experiments across various benchmarks, agent harnesses and LLMs, we isolate the effects of representation, outcome annota

作者 · Zhiyuan Jiang

在 arXiv 阅读

最新论文

全部论文 →

METAL LAB 最新报道