Towards harmonized regional style transfer and manipulation for facial images

12Citations
Citations of this article
14Readers
Mendeley users who have this article in their library.

Abstract

Regional facial image synthesis conditioned on a semantic mask has achieved great attention in the field of computational visual media. However, the appearances of different regions may be inconsistent with each other after performing regional editing. In this paper, we focus on harmonized regional style transfer for facial images. A multi-scale encoder is proposed for accurate style code extraction. The key part of our work is a multi-region style attention module. It adapts multiple regional style embeddings from a reference image to a target image, to generate a harmonious result. We also propose style mapping networks for multi-modal style synthesis. We further employ an invertible flow model which can serve as mapping network to fine-tune the style code by inverting the code to latent space. Experiments on three widely used face datasets were used to evaluate our model by transferring regional facial appearance between datasets. The results show that our model can reliably perform style transfer and multi-modal manipulation, generating output comparable to the state of the art. [Figure not available: see fulltext.]

Cite

CITATION STYLE

APA

Wang, C., Tang, F., Zhang, Y., Wu, T., & Dong, W. (2023). Towards harmonized regional style transfer and manipulation for facial images. Computational Visual Media, 9(2), 351–366. https://doi.org/10.1007/s41095-022-0284-6

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free