Abstract
Encoding an object essence in terms of self-similarities between its parts is becoming a popular strategy in Computer Vision. In this paper, a new similarity-based descriptor, dubbed Structural Similarity Cross-Covariance Tensor is proposed, aimed to encode relations among different regions of an image in terms of cross-covariance matrices. The latter are calculated between low-level feature vectors extracted from pairs of regions. The new descriptor retains the advantages of the widely used covariance matrix descriptors [1], extending their expressiveness from local similarities inside a region to structural similarities across multiple regions. The new descriptor, applied on top of HOG, is tested on object and scene classification tasks with three datasets. The proposed method always outclasses baseline HOG and yields significant improvement over a recently proposed self-similarity descriptor in the two most challenging datasets. © Springer-Verlag 2013.
Author supplied keywords
Cite
CITATION STYLE
San Biagio, M., Martelli, S., Crocco, M., Cristani, M., & Murino, V. (2013). Encoding classes of unaligned objects using structural similarity cross-covariance tensors. In Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) (Vol. 8258 LNCS, pp. 133–140). https://doi.org/10.1007/978-3-642-41822-8_17
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.