RGB-D scene recognition based on object-scene relation and semantics-preserving attention

1Citations
Citations of this article
35Readers
Mendeley users who have this article in their library.
Get full text

Abstract

Scene recognition is challenging due to intra-class diversity and inter-class similarity. Previous works recognize scenes either with global representations or with intermediate representations of objects. By contrast, we investigate more discriminative sequential representation of object-to-scene relations (SOSRs) for scene recognition. Particularly, we develop an Attention-Preserving Memory-Learning (APML) model, which enforces the Memory Network of the semantic domain to guide the Learning Network of the appearance domain in the learning procedure. Accordingly, we allocate semantics-preserving attention to different objects, which is more effective to seek the key encoded SOSR and discard the misleading encoded SOSR between objects and scene without requiring extra labeled data. Based on the proposed APML networks, we obtain the state-of-the-art results of RGB-D scene recognition on SUN RGB-D and NYUD2 datasets.

Cite

CITATION STYLE

APA

Guo, Y., & Liang, X. (2021). RGB-D scene recognition based on object-scene relation and semantics-preserving attention. In ICMR 2021 - Proceedings of the 2021 International Conference on Multimedia Retrieval (pp. 127–134). Association for Computing Machinery, Inc. https://doi.org/10.1145/3460426.3463603

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free