Does Multimodality Help Human and Machine for Translation and Image Captioning?

N/ACitations
Citations of this article
134Readers
Mendeley users who have this article in their library.

Abstract

This paper presents the systems developed by LIUM and CVC for the WMT16 Multimodal Machine Translation challenge. We explored various comparative methods, namely phrase-based systems and attentional recurrent neural networks models trained using monomodal or multimodal data. We also performed a human evaluation in order to estimate the usefulness of multimodal data for human machine translation and image description generation. Our systems obtained the best results for both tasks according to the automatic evaluation metrics BLEU and METEOR.

Cite

CITATION STYLE

APA

Caglayan, O., Aransa, W., Wang, Y., Masana, M., García-Martínez, M., Bougares, F., … Van De Weijer, J. (2016). Does Multimodality Help Human and Machine for Translation and Image Captioning? In Proceedings of the Annual Meeting of the Association for Computational Linguistics (Vol. 2, pp. 627–633). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/w16-2358

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free