Supervised and nonlinear alignment of two embedding spaces for dictionary induction in low resourced languages

2Citations
Citations of this article
86Readers
Mendeley users who have this article in their library.

Abstract

Enabling cross-lingual NLP tasks by leveraging multilingual word embedding has recently attracted much attention. An important motivation is to support lower resourced languages, however, most efforts focus on demonstrating the effectiveness of the techniques using embeddings derived from similar languages to English with large parallel content. In this study, we present a noise tolerant piecewise linear technique to learn a non-linear mapping between two monolingual word embedding vector spaces. We evaluate our approach on inferring bilingual dictionaries. We show that our technique outperforms the state of the art in lower resourced settings with an average improvement of 3.7% for precision @10 across 14 mostly low resourced languages.

Cite

CITATION STYLE

APA

Moshtaghi, M. (2019). Supervised and nonlinear alignment of two embedding spaces for dictionary induction in low resourced languages. In EMNLP-IJCNLP 2019 - 2019 Conference on Empirical Methods in Natural Language Processing and 9th International Joint Conference on Natural Language Processing, Proceedings of the Conference (pp. 823–832). Association for Computational Linguistics. https://doi.org/10.18653/v1/d19-1076

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free