Learnable Masks for Pose-Guided View Synthesis

1Citations
Citations of this article
3Readers
Mendeley users who have this article in their library.
Get full text

Abstract

Pose-guided human view synthesis uses a target pose to generate the appearance of a new view of a person. The input view and the target pose can be processed separately with UNet architectures that combine the results in a late fusion stage. UNet architectures link their encoder and decoder with skip connections that preserve the location of spatial features by injecting input information in the decoding process. However, direct skip connections may transfer irrelevant information to the decoder. We overcome this limitation with learnable masks for skip connections that encourage the decoder to use only relevant information from the encoder. We show that adding the proposed mask to UNet architectures improves the performance of view synthesis with only a slight increase in inference time.

Cite

CITATION STYLE

APA

Lakhal, M. I., Lanz, O., & Cavallaro, A. (2019). Learnable Masks for Pose-Guided View Synthesis. In Proceedings - International Conference on Image Processing, ICIP (Vol. 2019-September, pp. 1775–1779). IEEE Computer Society. https://doi.org/10.1109/ICIP.2019.8803134

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free