Image annotation incorporating low-rankness, tag and visual correlation and inhomogeneous errors

Yuqing Hou

Conference Proceedings

Image annotation incorporating low-rankness, tag and visual correlation and inhomogeneous errors

Hou Y

Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) (2015) 9474 71-81

DOI: 10.1007/978-3-319-27857-5_7

3Citations

1Readers

Get full text

Abstract

Tag-based image retrieval (TBIR) has drawn much attention in recent years due to the explosive amount of digital images and crowdsourcing tags. However, TBIR is still suffering from the incomplete and inaccurate tags provided by users, posing a great challenge for tag-based image management applications. In this work, we propose a novel method for image annotation, incorporating several priors: Low- Rankness, Tag and Visual Correlation and Inhomogeneous Errors. Highly representative CNN feature vectors are adopted to model the tag-visual correlation and narrow the semantic gap. And we extract word vectors for tags to measure similarity between tags in the semantic level, which is more accurate than traditional frequency-based or graph-based methods. We utilize the Accelerated Proximal Gradient (APG) method to solve our model efficiently. Extensive experiments conducted on multiple benchmark datasets demonstrate the effectiveness and robustness of the proposed method.

Cite

CITATION STYLE

APA

Hou, Y. (2015). Image annotation incorporating low-rankness, tag and visual correlation and inhomogeneous errors. In Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) (Vol. 9474, pp. 71–81). Springer Verlag. https://doi.org/10.1007/978-3-319-27857-5_7

Image annotation incorporating low-rankness, tag and visual correlation and inhomogeneous errors

Abstract

Cite

Register to see more suggestions