Imitator: Personalized Speech-driven 3D Facial Animation

编辑：映维 | 分类：CV / XR | 2023年1月18日

Note: We don't have the ability to review paper

PubDate: Dec 2022

Teams: Max Planck Institute for Intelligent Systems;Microsoft Mixed Reality & AI Lab

Writers: Balamurugan Thambiraja, Ikhsanul Habibie, Sadegh Aliakbarian, Darren Cosker, Christian Theobalt, Justus Thies

PDF: Imitator: Personalized Speech-driven 3D Facial Animation

Imitator: Personalized Speech-driven 3D Facial Animation

Abstract

Speech-driven 3D facial animation has been widely explored, with applications in gaming, character animation, virtual reality, and telepresence systems. State-of-the-art methods deform the face topology of the target actor to sync the input audio without considering the identity-specific speaking style and facial idiosyncrasies of the target actor, thus, resulting in unrealistic and inaccurate lip movements. To address this, we present Imitator, a speech-driven facial expression synthesis method, which learns identity-specific details from a short input video and produces novel facial expressions matching the identity-specific speaking style and facial idiosyncrasies of the target actor. Specifically, we train a style-agnostic transformer on a large facial expression dataset which we use as a prior for audio-driven facial expressions. Based on this prior, we optimize for identity-specific speaking style based on a short reference video. To train the prior, we introduce a novel loss function based on detected bilabial consonants to ensure plausible lip closures and consequently improve the realism of the generated expressions. Through detailed experiments and a user study, we show that our approach produces temporally coherent facial expressions from input audio while preserving the speaking style of the target actors.

本文链接：https://paper.nweon.com/13838

Imitator: Personalized Speech-driven 3D Facial Animation

您可能还喜欢...

最新AR/VR行业分享

最新AR/VR专利

最新AR/VR行业招聘

Imitator: Personalized Speech-driven 3D Facial Animation

您可能还喜欢...

Transitioning360: Content-aware NFoV Virtual Camera Paths for 360° Video Playback

Comparative Study of Latent-Sensitive Processing of Heterogeneous Data in an Experimental Platform for 3D Video Holographic Communication

Real-Time Marker-Based Finger Tracking with Neural Networks

最新AR/VR行业分享

最新AR/VR专利

最新AR/VR行业招聘