Analyzing Architectures for Neural Machine Translation using Low Computational Resources

  • Mandke A
  • Litake O
  • Kadam D
N/ACitations
Citations of this article
6Readers
Mendeley users who have this article in their library.

Abstract

With the recent developments in the field of Natural Language Processing, there has been a rise in the use of different architectures for Neural Machine Translation. Transformer architectures are used to achieve state-of-the-art accuracy, but they are very computationally expensive to train. Everyone cannot have such setups consisting of high-end GPUs and other resources. We train our models on low computational resources and investigate the results. As expected, transformers outperformed other architectures, but there were some surprising results. Transformers consisting of more encoders and decoders took more time to train but had fewer BLEU scores. LSTM performed well in the experiment and took comparatively less time to train than transformers, making it suitable to use in situations having time constraints.

Cite

CITATION STYLE

APA

Mandke, A., Litake, O., & Kadam, D. (2021). Analyzing Architectures for Neural Machine Translation using Low Computational Resources. International Journal on Natural Language Computing, 10(5), 9–16. https://doi.org/10.5121/ijnlc.2021.10502

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free