Leveraging pre-trained embeddings forwelsh taggers

7Citations
Citations of this article
76Readers
Mendeley users who have this article in their library.

Abstract

While the application of word embedding models to downstream Natural Language Processing (NLP) tasks has been shown to be successful, the benefits for low-resource languages is somewhat limited due to lack of adequate data for training the models. However, NLP research efforts for low-resource languages have focused on constantly seeking ways to harness pre-trained models to improve the performance of NLP systems built to process these languages without the need to reinvent the wheel. One such language is Welsh and therefore, in this paper, we present the results of our experiments on learning a simple multi-task neural network model for partof- speech and semantic tagging for Welsh using a pre-trained embedding model from Fast- Text. Our model's performance was compared with those of the existing rule-based standalone taggers for part-of-speech and semantic taggers. Despite its simplicity and capacity to perform both tasks simultaneously, our tagger compared very well with the existing taggers.

Cite

CITATION STYLE

APA

Ezeani, I., Piao, S., Neale, S., Rayson, P., & Knight, D. (2019). Leveraging pre-trained embeddings forwelsh taggers. In ACL 2019 - 4th Workshop on Representation Learning for NLP, RepL4NLP 2019 - Proceedings of the Workshop (pp. 270–280). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/w19-4332

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free