Subject
2 entries
Talking Head
Bookmarks
Speech Driven Talking Head Generation via Attentional Landmarks Based Representation
This paper introduces an attentional landmark-based representation for generating realistic talking head video from speech audio, using facial landmarks as a compact intermediate representation that bridges audio and visual domains. The approach decouples appearance generation from motion modeling, improving generalization across identities.
AD-NeRF: Audio Driven Neural Radiance Fields for Talking Head Synthesis
AD-NeRF generates photorealistic talking-head video directly from audio using neural radiance fields, bypassing the 2D landmarks or 3D face model intermediaries used by prior methods. By conditioning an implicit neural function on audio features and rendering via volume rendering, it achieves both head and upper body generation with free-viewpoint control.
