Spoken language identification using i-vectors, x-vectors, plda and logistic regression

Ahmad Iqbal Abdurrahman; Amalia Zahra

Journal ArticleOPEN ACCESS

Spoken language identification using i-vectors, x-vectors, plda and logistic regression

Bulletin of Electrical Engineering and Informatics (2021) 10(4) 2237-2244

DOI: 10.11591/EEI.V10I4.2893

14Citations

16Readers

Abstract

In this paper, i-vector and x-vector is used to extract the features from speech signal from local Indonesia languages, namely Javanese, Sundanese and Minang languages to help classifier identify the language spoken by the speaker. Probabilistic linear discriminant analysis (PLDA) are used as the baseline classifier and logistic regression technique are used because of prior studies showing logistic regression has better performance than PLDA for classifying speech data. Once these features are extracted. The feature is going to be classified using the classifier mentioned before. In the experiment, we tried to segment the test data to three segment such as 3, 10, and 30 seconds. This study is expanded by testing multiple parameters on the i-vector and x-vector method then comparing PLDA and logistic regression performance as its classifier. The x-vector has better score than i-vector for every segmented data while using PLDA as its classifier, except where the i-vector and x-vector is using logistic regression, i-vector still has better accuracy compared to x-vector.

Author supplied keywords

Cite

CITATION STYLE

APA

Abdurrahman, A. I., & Zahra, A. (2021). Spoken language identification using i-vectors, x-vectors, plda and logistic regression. Bulletin of Electrical Engineering and Informatics, 10(4), 2237–2244. https://doi.org/10.11591/EEI.V10I4.2893

Spoken language identification using i-vectors, x-vectors, plda and logistic regression

Abstract

Author supplied keywords

Cite

Register to see more suggestions