arXiv cs.AIAugust 18, 2026
CETalk: Continuous Valence-Arousal Control for Audio-Driven 3D Talking Head Generation
Excerpt
arXiv:2608.15110v1 Announce Type: cross Abstract: Emotional 3D talking head generation aims to synthesize expressive facial animations with accurate lip synchronization. However, existing methods often rely on discrete emotion categories, which fail to capture the continuous evolution of affect. They also overlook the temporal frequency mismatch between audio articulation and emotional expression. In this paper, we propose CETalk, an audio-driven 3D facial animation framework conditioned on cont