Skip to content

Author

Menglin Han

We have 3 of 8 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

Vorch-Human: Unified Multi-Task Human-Centric Generation via Long-Horizon Continuation

Human-centric audio-visual generation spans several closely related tasks: animating a person from driving speech, jointly generating speech and video from a voice reference, and synthesizing a scene from paired appearance and voice references. Existing systems commonly solve these tasks with separate models, even thou...

Yang Ding, Hao-Ran Yu, Xin Ma et al. · 0 citations
Preprint Aug 2026

Vorch-Omni: Multi-Task Orchestration of Sight and Sound

Vorch-Omni is presented, a unified multi-task framework for audio-visual synthesis based on an arbitrary-condition-to-arbitrary-output formulation that supports over 10 tasks, including text-to-video, text-to-audio-video, image- and reference-conditioned generation, temporal extension, audio-driven generation, video tr...

Vorch Team, Xiaoyu Chen, Yang Ding et al. · 0 citations
Preprint Aug 2026

Vorch-Streamer: Extending Human Audio-Visual Generation to Real-Time Long-Form Streaming

Vorch-Streamer, a post-training framework that addresses real-time long-form Text-to-Audio-Video (T2AV) streaming and enables real-time long-form avatar audio-video streaming, and bounded causal context and four-step denoising are presented.

Meng-Lin Han, Yang Ding, Yu-Lei Lu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.