Learning to describe E-commerce images from noisy online data

Takuya Yashima; Naoaki Okazaki; Kentaro Inui; Kota Yamaguchi; Takayuki Okatani

Conference Proceedings

Learning to describe E-commerce images from noisy online data

Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) (2017) 10115 LNCS 85-100

DOI: 10.1007/978-3-319-54193-8_6

4Citations

26Readers

Get full text

Abstract

Recent study shows successful results in generating a proper language description for the given image, where the focus is on detecting and describing the contextual relationship in the image, such as the kind of object, relationship between two objects, or the action. In this paper, we turn our attention to more subjective components of descriptions that contain rich expressions to modify objects – namely attribute expressions. We start by collecting a large amount of product images from the online market site Etsy, and consider learning a language generation model using a popular combination of a convolutional neural network (CNN) and a recurrent neural network (RNN). Our Etsy dataset contains unique noise characteristics often arising in the online market. We first apply natural language processing techniques to extract highquality, learnable examples in the real-world noisy data. We learn a generation model from product images with associated title descriptions, and examine how e-commerce specific meta-data and fine-tuning improve the generated expression. The experimental results suggest that we are able to learn from the noisy online data and produce a product description that is closer to a man-made description with possibly subjective attribute expressions.

Cite

CITATION STYLE

APA

Yashima, T., Okazaki, N., Inui, K., Yamaguchi, K., & Okatani, T. (2017). Learning to describe E-commerce images from noisy online data. In Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) (Vol. 10115 LNCS, pp. 85–100). Springer Verlag. https://doi.org/10.1007/978-3-319-54193-8_6

Learning to describe E-commerce images from noisy online data

Abstract

Cite

Register to see more suggestions