CETalk: Continuous Valence-Arousal Control for Audio-Driven 3D Talking Head Generation
Emotional 3D talking head generation aims to synthesize expressive facial animations with accurate lip synchronization. However, existing methods often rely on discrete emotion categories, which fail to capture the continuous evolution of affect. They also overlook the temporal frequency mismatch between audio articula...