One email each morning — yesterday's AI, sortedGet it in your inbox

METAL LAB

Demystifying Agent Skills: Why They Work-Until They Don't

arXiv:2608.140362026-08-13

Skills have emerged as a practical and effective approach for enhancing LLM agents at inference time through structured packages of knowledge. However, existing evaluations largely measure whether skills improve aggregated task success, leaving a more fundamental question underexplored: \textbf{When do skills help, why do they work, and where do they fail?} Through controlled experiments across various benchmarks, agent harnesses and LLMs, we isolate the effects of representation, outcome annota

Authors · Zhiyuan Jiang

Read on arXiv

Latest papers

All papers →

Latest from METAL LAB