Improving the performance of vietnamese–korean neural machine translation with contextual embedding

Van Hai Vu; Quang Phuoc Nguyen; Ebipatei Victoria Tunyan; Cheol Young Ock

Journal ArticleOPEN ACCESS

Improving the performance of vietnamese–korean neural machine translation with contextual embedding

Applied Sciences (Switzerland) (2021) 11(23)

DOI: 10.3390/app112311119

3Citations

10Readers

Abstract

With the recent evolution of deep learning, machine translation (MT) models and systems are being steadily improved. However, research on MT in low-resource languages such as Vietnamese and Korean is still very limited. In recent years, a state-of-the-art context-based embedding model introduced by Google, bidirectional encoder representations for transformers (BERT), has begun to appear in the neural MT (NMT) models in different ways to enhance the accuracy of MT systems. The BERT model for Vietnamese has been developed and significantly improved in natural language processing (NLP) tasks, such as part-of-speech (POS), named-entity recognition, dependency parsing, and natural language inference. Our research experimented with applying the Vietnamese BERT model to provide POS tagging and morphological analysis (MA) for Vietnamese sentences„ and applying word-sense disambiguation (WSD) for Korean sentences in our Vietnamese–Korean bilingual corpus. In the Vietnamese–Korean NMT system, with contextual embedding, the BERT model for Vietnamese is concurrently connected to both encoder layers and decoder layers in the NMT model. Experimental results assessed through BLEU, METEOR, and TER metrics show that contextual embedding significantly improves the quality of Vietnamese–Korean NMT.

Author supplied keywords

Cite

CITATION STYLE

APA

Vu, V. H., Nguyen, Q. P., Tunyan, E. V., & Ock, C. Y. (2021). Improving the performance of vietnamese–korean neural machine translation with contextual embedding. Applied Sciences (Switzerland), 11(23). https://doi.org/10.3390/app112311119

Improving the performance of vietnamese–korean neural machine translation with contextual embedding

Abstract

Author supplied keywords

Cite

Register to see more suggestions