One email each morning — yesterday's AI, sortedGet it in your inbox

METAL LAB

CoinVE-200K: A Large-Scale High-Quality Dataset for Compositional Instruction-Guided Video Editing

arXiv:2608.175662026-08-17

The quality and diversity of instruction-based video editing datasets are steadily improving, yet existing datasets mainly focus on single editing operations and fall short in supporting compositional instruction-guided video editing. In particular, multiple editing intents must be jointly understood and faithfully executed within the same video. To address this issue, we introduce CoinVE-200K, a large-scale, high-quality dataset for Compositional Instruction-Guided Video Editing. CoinVE-200K co

Authors · Fuchen Long

Read on arXiv

Latest papers

All papers →

Latest from METAL LAB