Facial action unit intensity estimation via semantic correspondence learning with dynamic graph convolution

45Citations
Citations of this article
38Readers
Mendeley users who have this article in their library.
Get full text

Abstract

The intensity estimation of facial action units (AUs) is challenging due to subtle changes in the person’s facial appearance. Previous approaches mainly rely on probabilistic models or predefined rules for modeling co-occurrence relationships among AUs, leading to limited generalization. In contrast, we present a new learning framework that automatically learns the latent relationships of AUs via establishing semantic correspondences between feature maps. In the heatmap regression-based network, feature maps preserve rich semantic information associated with AU intensities and locations. Moreover, the AU co-occurring pattern can be reflected by activating a set of feature channels, where each channel encodes a specific visual pattern of AU. This motivates us to model the correlation among feature channels, which implicitly represents the co-occurrence relationship of AU intensity levels. Specifically, we introduce a semantic correspondence convolution (SCC) module to dynamically compute the correspondences from deep and low resolution feature maps, and thus enhancing the discriminability of features. The experimental results demonstrate the effectiveness and the superior performance of our method on two benchmark datasets.

Cite

CITATION STYLE

APA

Fan, Y., Lam, J. C. K., & Li, V. O. K. (2020). Facial action unit intensity estimation via semantic correspondence learning with dynamic graph convolution. In AAAI 2020 - 34th AAAI Conference on Artificial Intelligence (pp. 12701–12708). AAAI press. https://doi.org/10.1609/aaai.v34i07.6963

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free