CVPR 2023 Text Guided Video Editing Competition

Bai, Jinbin; Cheng, Xu; Dong, Zhen; Gao, Difei; He, Rui; Hu, Feng; Hu, Junhua; Huang, Hai; Huang, Zuwei; Iandola, Forrest; Keutzer, Kurt; Li, Xiuyu; Li, Youzeng; Shou, Mike Zheng; Singh, Aishani; Sun, Yuanxi; Tang, Jie; Wu, Jay Zhangjie; Xiang, Xiaoyu; Zhu, Hanyu

CVPR 2023 Text Guided Video Editing Competition

Authors: Jinbin Bai
Xu Cheng
Zhen Dong
Difei Gao
Rui He
Feng Hu
Junhua Hu
Hai Huang
Zuwei Huang
Forrest Iandola
Kurt Keutzer
Xiuyu Li
Youzeng Li
Mike Zheng Shou
Aishani Singh
Yuanxi Sun
Jie Tang
Jay Zhangjie Wu
Xiaoyu Xiang
Hanyu Zhu
Publication date: 24 October 2023
Publisher

Abstract

Humans watch more than a billion hours of video per day. Most of this video was edited manually, which is a tedious process. However, AI-enabled video-generation and video-editing is on the rise. Building on text-to-image models like Stable Diffusion and Imagen, generative AI has improved dramatically on video tasks. But it's hard to evaluate progress in these video tasks because there is no standard benchmark. So, we propose a new dataset for text-guided video editing (TGVE), and we run a competition at CVPR to evaluate models on our TGVE dataset. In this paper we present a retrospective on the competition and describe the winning method. The competition dataset is available at https://sites.google.com/view/loveucvpr23/track4.Comment: Project page: https://sites.google.com/view/loveucvpr23/track

Similar works

Full text

Available Versions

arXiv.org e-Print Archive

oai:arXiv.org:2310.16003

Last time updated on 16/01/2024