Abstract
Despite the successful use of local image features for large-scale object recognition, they are not effective in recognizing book spines on bookshelves. This is because some book spines contain only text components that do not yield distinguishing image features. To overcome this issue, we develop a new approach that combines a text-based spine recognition pipeline with an image feature-based spine recognition pipeline. The text within the book spine image is recognized and used as keywords to search a book spine text database. The image features of the book spine image are searched through a book spine image database. The search results of the two approaches are then carefully combined to form the final result. We implement the proposed hybrid book recognition pipeline used in a book inventory management system, and conduct extensive experiments to evaluate its performance. The experimental results show that while text-based or image feature-based systems only achieve a recall of ∼72%, the proposed hybrid system achieves a recall of ∼91%.
Author supplied keywords
Cite
CITATION STYLE
Tsai, S. S., Chen, D., Chen, H., Hsu, C. H., Kim, K. H., Singh, J. P., & Girod, B. (2011). Combining image and text features: A hybrid approach to mobile book spine recognition. In MM’11 - Proceedings of the 2011 ACM Multimedia Conference and Co-Located Workshops (pp. 1029–1032). https://doi.org/10.1145/2072298.2071930
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.