每天早上一封邮件,把昨天的 AI 梳理好订阅邮件

METAL LAB

SemComp-Bench: Benchmarking Semantic Task Completion in Video Generation

arXiv:2608.174262026-08-17

We introduce Semantic Task Completion Video Generation, an outcome-oriented video generation task. Under this formulation, success requires both achievement of the intended outcome and semantic grounding. Semantic grounding characterizes the correspondence between the reference image and the generated outcome in terms of high-level semantics relevant to the task. Evaluation focuses on the generated outcome and requires neither the presentation of a complete sequence of intermediate task steps no

作者 · Keyu Tu

在 arXiv 阅读

最新论文

全部论文 →

METAL LAB 最新报道