ManVatar : Fast 3D Head Avatar Reconstruction Using Motion-Aware Neural Voxels

International Conference on Computer Graphics and Interactive Techniques (SIGGRAPH), 2022

23 November 2022

Yuelang Xu

Lizhen Wang

Xiaochen Zhao

Hongwen Zhang

Yebin Liu

3DH

ArXiv (abs)PDF HTML

Abstract

With NeRF widely used for facial reenactment, recent methods can recover photo-realistic 3D head avatar from just a monocular video. Unfortunately, the training process of the NeRF-based methods is quite time-consuming, as MLP used in the NeRF-based methods is inefficient and requires too many iterations to converge. To overcome this problem, we propose ManVatar, a fast 3D head avatar reconstruction method using Motion-Aware Neural Voxels. ManVatar is the first to decouple expression motion from canonical appearance for head avatar, and model the expression motion by neural voxels. In particular, the motion-aware neural voxels is generated from the weighted concatenation of multiple 4D tensors. The 4D tensors semantically correspond one-to-one with 3DMM expression bases and share the same weights as 3DMM expression coefficients. Benefiting from our novel representation, the proposed ManVatar can recover photo-realistic head avatars in just 5 minutes (implemented with pure PyTorch), which is significantly faster than the state-of-the-art facial reenactment methods.

View on arXiv

Comments on this paper