PATMAT: Person Aware Tuning of Mask-Aware Transformer for Face Inpainting

Motamed, Saman; Xu, Jian**; Wu, Chen Henry; De la Torre, Fernando

Computer Science > Computer Vision and Pattern Recognition

arXiv:2304.06107 (cs)

[Submitted on 12 Apr 2023]

Title:PATMAT: Person Aware Tuning of Mask-Aware Transformer for Face Inpainting

Authors:Saman Motamed, Jian** Xu, Chen Henry Wu, Fernando De la Torre

View PDF

Abstract:Generative models such as StyleGAN2 and Stable Diffusion have achieved state-of-the-art performance in computer vision tasks such as image synthesis, inpainting, and de-noising. However, current generative models for face inpainting often fail to preserve fine facial details and the identity of the person, despite creating aesthetically convincing image structures and textures. In this work, we propose Person Aware Tuning (PAT) of Mask-Aware Transformer (MAT) for face inpainting, which addresses this issue. Our proposed method, PATMAT, effectively preserves identity by incorporating reference images of a subject and fine-tuning a MAT architecture trained on faces. By using ~40 reference images, PATMAT creates anchor points in MAT's style module, and tunes the model using the fixed anchors to adapt the model to a new face identity. Moreover, PATMAT's use of multiple images per anchor during training allows the model to use fewer reference images than competing methods. We demonstrate that PATMAT outperforms state-of-the-art models in terms of image quality, the preservation of person-specific details, and the identity of the subject. Our results suggest that PATMAT can be a promising approach for improving the quality of personalized face inpainting.

Subjects:	Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
Cite as:	arXiv:2304.06107 [cs.CV]
	(or arXiv:2304.06107v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2304.06107

Submission history

From: Saman Motamed [view email]
[v1] Wed, 12 Apr 2023 18:46:37 UTC (26,517 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:PATMAT: Person Aware Tuning of Mask-Aware Transformer for Face Inpainting

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:PATMAT: Person Aware Tuning of Mask-Aware Transformer for Face Inpainting

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators