Audio-driven avatar generation requires realistic lip-sync, expressive motion, and real-time streaming. Recent work achieves the latter via self-forcing with Distribution Matching Distillation (DMD), but this paradigm suffers from a critical failure that has not been systematically characterized: dynamic collapse, wher...
Yubo Huang, Si-Rui Zhao, Xinchen Yao et al.· 1 citation· ⚡1
Multimodal emotion recognition based on complementary physiological signals such as electroencephalogram (EEG) and eye movements can effectively reflect human emotional states, demonstrating significant potential in fields such as rehabilitation monitoring and driving safety. However, existing multimodal emotion recogn...
Xin-Hui Li, Hao-Yuan Chen, Minchao Wu et al.· ACM Transactions on Autonomo...· 0 citations
DFCS is proposed, a training-free, trigger-agnostic method that clusters fixed pretrained features into one region per poisoning slot and selects the centroid-nearest sample from each region and supports distributional feature coverage as an effective selection principle for low-budget dirty-label backdoor attacks.
Yi Yang, Xiaoke Chen, Jin-Yang Huang et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.