Automatic Bilingual Phrase Dictionary Construction from GIZA++ Output

0Citations
Citations of this article
29Readers
Mendeley users who have this article in their library.

Abstract

Modern encoder-decoder based neural machine translation (NMT) models are normally trained on parallel sentences. Hence, they give best results when translating full sentences rather than sentence parts. Thereby, the task of translating commonly used phrases, which often arises for language learners, is not addressed by NMT models. While for high-resourced language pairs human-built phrase dictionaries exist, less-resourced pairs do not have them. We suggest an approach for building such dictionary automatically based on the GIZA++ output and show that it works significantly better than translating phrases with a sentences-trained NMT system.

Cite

CITATION STYLE

APA

Khusainova, A., Romanov, V., & Khan, A. (2022). Automatic Bilingual Phrase Dictionary Construction from GIZA++ Output. In LREC 2022 Workshop - Language Resources and Evaluation Conference, 18th Workshop on Multiword Expressions, MWE 2022 - Proceedings (pp. 81–88). European Language Resources Association (ELRA). https://doi.org/10.28995/2075-7182-2022-21-1068-1077

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free