Fully automatic camera movement directly affects the art quality of dance expressiveness, especially in terms of visual expression, as well as choreography and music. Current studies mainly focus on synthesizing camera movements conditioned on dance and music, but they overlook the camera movement style, which is essen...
Xiao-Ying Huang, Sanyi Zhang, Xi-Rui Wang et al.· Proceedings of the Thirty-Fi...· 0 citations
Recent audio generation models can synthesize high-fidelity speech, environmental sound, singing voice, and music, creating new risks for multimedia trust. Existing audio deepfake detection (ADD) benchmarks remain predominantly speech-centric and often underrepresent realistic channel variation and diverse audio types....
Yuan-Kun Xie, Hao-Nan Cheng, Jiayi Zhou et al.· 1 citation
This paper summarizes the ACM Multimedia 2026 AT-ADD Grand Challenge on all-type audio deepfake detection. AT-ADD contains two tracks: robust speech deepfake detection under realistic acoustic and channel variations, and type-agnostic detection over speech, environmental sound, singing voice, and music. We describe the...
Yuan-Kun Xie, Hao-Nan Cheng, Jiayi Zhou et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.