Creating animatable and relightable human avatars from multi-view images remains challenging, as pose-dependent deformation, materials, and light visibility are tightly coupled in images. In this paper, we present ARS-Avatar, a novel method using surfel representation for high-quality, animatable, and relightable human...
Jia-Teng Liu, Hao Gao, Jun-Xin Sun et al.· 0 citations
Current music-to-dance generation methods mainly rely on musical features, limiting precise control over generated movements. In particular, most existing methods with control mechanisms do not support example-based control, in which a user provides a reference motion sequence and the generated dances follow its fine-g...
Meng-Qi Liu, Hao Gao, Haolun Li et al.· IEEE Transactions on Visuali...· 0 citations
A self-distillation framework that leverages large-scale 2D talking videos to pre-train a specialized speech encoder that aligns speech representations with expressive visual dynamics, allowing the encoder to extract features highly correlated with facial and head motions directly from audio.
Jiu-Cheng Xie, Ji-Wang Zheng, Yongkang Xia et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.