매일 아침, 어제의 AI를 한 통으로 정리해 보내드립니다메일로 받아보기

METAL LAB

WebGrader: Training LLMs for Web Development with Self-Evolving Programmatic Grader

arXiv:2608.064742026-08-10

arXiv:2608.06474v1 Announce Type: new Abstract: Large language models increasingly generate complete websites from natural-language descriptions, and reinforcement learning has become a central approach to closing their remaining functional gap. This training regime is bottlenecked by reward design. Hand-authored browser scripts are executable yet costly to write for open-ended requirements, while VLM and GUI-agent graders scale but may issue verdicts before observing the decisive state. We prop

저자 · Boshui Chen, Huiping Liu, Shaolei Zhang

arXiv에서 원문 보기