Worse WER, but better BLEU? Leveraging word embedding as intermediate in multitask end-to-end speech translation

17Citations
Citations of this article
107Readers
Mendeley users who have this article in their library.

Abstract

Speech translation (ST) aims to learn transformations from speech in the source language to the text in the target language. Previous works show that multitask learning improves the ST performance, in which the recognition decoder generates the text of the source language, and the translation decoder obtains the final translations based on the output of the recognition decoder. Because whether the output of the recognition decoder has the correct semantics is more critical than its accuracy, we propose to improve the multitask ST model by utilizing word embedding as the intermediate.

Cite

CITATION STYLE

APA

Chuang, S. P., Sung, T. W., Liu, A. H., & Lee, H. Y. (2020). Worse WER, but better BLEU? Leveraging word embedding as intermediate in multitask end-to-end speech translation. In Proceedings of the Annual Meeting of the Association for Computational Linguistics (pp. 5998–6003). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/2020.acl-main.533

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free