Lune

CVPR2021Top-tier venue

Audio-Driven Emotional Video Portraits

Xinya Ji, Hang Zhou, Kaisiyuan Wang, Wayne Wu, Chen Change Loy, Xun Cao, Feng Xu

2021Year
46Top-tier citations

Abstract

Target ID-1 Sad Emotion Interpolation (a) (b) Happy Sad Figure 1: Audio-Driven Emotional Video Portraits. Given an audio clip and a target video, our Emotional Video Portraits (EVP) approach is capable of generating emotion-controllable talking portraits and change the emotion of them smoothly by interpolating at the latent space. (a) Generated video portraits with the same speech content but different emotions (i.e., contempt and sad). (b) Linear interpolation of the learned latent representation of emotions from sad to happy.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 59a3e154-c023-49bb-b4af-31f5eb20328d

Cited by top-tier papers46

Ask how each one uses it

Builds on2

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines