Single Image, Any Face: Generalisable 3D Face Generation

Wang, Wenqing; Yang, Haosen; Kittler, Josef; Zhu, Xiatian

Computer Science > Computer Vision and Pattern Recognition

arXiv:2409.16990 (cs)

[Submitted on 25 Sep 2024 (v1), last revised 12 Mar 2025 (this version, v2)]

Title:Single Image, Any Face: Generalisable 3D Face Generation

Authors:Wenqing Wang, Haosen Yang, Josef Kittler, Xiatian Zhu

View PDF HTML (experimental)

Abstract:The creation of 3D human face avatars from a single unconstrained image is a fundamental task that underlies numerous real-world vision and graphics applications. Despite the significant progress made in generative models, existing methods are either less suited in design for human faces or fail to generalise from the restrictive training domain to unconstrained facial images. To address these limitations, we propose a novel model, Gen3D-Face, which generates 3D human faces with unconstrained single image input within a multi-view consistent diffusion framework. Given a specific input image, our model first produces multi-view images, followed by neural surface construction. To incorporate face geometry information in a generalisable manner, we utilise input-conditioned mesh estimation instead of ground-truth mesh along with synthetic multi-view training data. Importantly, we introduce a multi-view joint generation scheme to enhance appearance consistency among different views. To the best of our knowledge, this is the first attempt and benchmark for creating photorealistic 3D human face avatars from single images for generic human subject across domains. Extensive experiments demonstrate the superiority of our method over previous alternatives for out-of-domain singe image 3D face generation and top competition for in-domain setting.

Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2409.16990 [cs.CV]
	(or arXiv:2409.16990v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2409.16990

Submission history

From: Wenqing Wang [view email]
[v1] Wed, 25 Sep 2024 14:56:37 UTC (4,497 KB)
[v2] Wed, 12 Mar 2025 13:34:01 UTC (5,602 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Single Image, Any Face: Generalisable 3D Face Generation

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Single Image, Any Face: Generalisable 3D Face Generation

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators