PAPER / ARXIV:2609.11638
Jintao Zhang , Kai Jiang , Jintao Chen , Xu Wang , Deyuan Liu , Jungang Li , Dechuang Chen , Ming Lin , Jingjiang Zhou , Haopeng Jin , Qi Jia , Xiaohang Wang , Yaole Wang , Zhanqiang Zhang , Ran Li , Zhengkun Huang , Shuyue Xiong , Yuji Wang , Zikun Dai , Hui He , Yang Luo , Mang Ning , Weiqi Feng , Chengyang Ye , Xinyue Lin , Min Zhao , Hongzhou Zhu , Hengkai Tan , Zeyuan Wang , Chendong Xiang , Kaiwen Zheng , Zhijie Deng , Fan Bao , Jianfei Chen , Jun Zhu
RESUMO
We present Vidu S2, which comprises Vidu S2-Avatar, a real-time interactive digital-character model, and Vidu S2-Editing, a real-time video editing model. Moreover, we explore the feasibility of real-time spatial video generation for both Vidu S2-Avatar and Vidu S2-Editing. Compared with Vidu S1, Vidu S2-Avatar supports real-time 720p video generation, generation with dynamic references that can be updated at any moment, and stronger instruction following, such as dancing. Vidu S2-Editing supports editing a video stream in real time, including style rendering, clothing replacement, character replacement, and background replacement. Experiments show that Vidu S2 outperforms all baselines. A playable online demo is available at this https URL .
NO MESMO MAPA