Spelling correction of non-word errors in Uyghur-Chinese machine translation

Rui Dong; Yating Yang; Tonghai Jiang

Journal ArticleOPEN ACCESS

Spelling correction of non-word errors in Uyghur-Chinese machine translation

Information (Switzerland) (2019) 10(6)

DOI: 10.3390/info10060202

4Citations

26Readers

Abstract

This research was conducted to solve the out-of-vocabulary problem caused by Uyghur spelling errors in Uyghur-Chinese machine translation, so as to improve the quality of Uyghur-Chinese machine translation. This paper assesses three spelling correction methods based on machine translation: 1. Using a Bilingual Evaluation Understudy (BLEU) score; 2. Using a Chinese language model; 3. Using a bilingual language model. The best results were achieved in both the spelling correction task and the machine translation task by using the BLEU score for spelling correction. A maximum F1 score of 0.72 was reached for spelling correction, and the translation result increased the BLEU score by 1.97 points, relative to the baseline system. However, the method of using a BLEU score for spelling correction requires the support of a bilingual parallel corpus, which is a supervised method that can be used in corpus pre-processing. Unsupervised spelling correction can be performed by using either a Chinese language model or a bilingual language model. These two methods can be easily extended to other languages, such as Arabic.

Author supplied keywords

Cite

CITATION STYLE

APA

Dong, R., Yang, Y., & Jiang, T. (2019). Spelling correction of non-word errors in Uyghur-Chinese machine translation. Information (Switzerland), 10(6). https://doi.org/10.3390/info10060202

Spelling correction of non-word errors in Uyghur-Chinese machine translation

Abstract

Author supplied keywords

Cite

Register to see more suggestions