Subject
1 entry
Speech Synthesis
Bookmarks
Speech Driven Talking Head Generation via Attentional Landmarks Based Representation
This paper introduces an attentional landmark-based representation for generating realistic talking head video from speech audio, using facial landmarks as a compact intermediate representation that bridges audio and visual domains. The approach decouples appearance generation from motion modeling, improving generalization across identities.
