Skip to content

Author

Yaowei Wang

3 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

Vorch-Human: Unified Multi-Task Human-Centric Generation via Long-Horizon Continuation

Human-centric audio-visual generation spans several closely related tasks: animating a person from driving speech, jointly generating speech and video from a voice reference, and synthesizing a scene from paired appearance and voice references. Existing systems commonly solve these tasks with separate models, even thou...

Yang Ding, Hao-Ran Yu, Xin Ma et al. · 0 citations
Preprint Aug 2026

Vorch-Omni: Multi-Task Orchestration of Sight and Sound

Vorch-Omni is presented, a unified multi-task framework for audio-visual synthesis based on an arbitrary-condition-to-arbitrary-output formulation that supports over 10 tasks, including text-to-video, text-to-audio-video, image- and reference-conditioned generation, temporal extension, audio-driven generation, video tr...

Vorch Team, Xiaoyu Chen, Yang Ding et al. · 0 citations
Preprint Aug 2026

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation

Vorch-IR is presented, a unified framework that supports single- and dual-person identity replacement, with optional background replacement, in a single model, and an automatic data construction pipeline that synthesizes paired supervision for all four editing settings is developed.

Yaowei Wang, Xiaoyu Chen, Xin Ma et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.