Study on Tibetan Word Vector based on Word2vec

2Citations
Citations of this article
7Readers
Mendeley users who have this article in their library.

This article is free to access.

Abstract

This paper uses Word2vec to study Tibetan word vector. Word2vec is optimized by two methods: Hierarchical Softmax and Negative Sampling in CBOW and Skip-gram models. Through the training of neural network, the words in Tibetan sentences are converted into vector form. Word2vec transforms the Tibetan text content processing into a simple vector space operation, calculates the similarity in the vector space, and then obtains the semantic similarity of the text, providing an accurate word vector for the training of the language model.

Cite

CITATION STYLE

APA

Yang, N., Li, G., Ding, H., & Gong, C. (2019). Study on Tibetan Word Vector based on Word2vec. In Journal of Physics: Conference Series (Vol. 1187). Institute of Physics Publishing. https://doi.org/10.1088/1742-6596/1187/5/052074

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free