Online speaker clustering using incremental learning of an ergodic hidden Markov model

4Citations
Citations of this article
5Readers
Mendeley users who have this article in their library.

Abstract

A Novel online speaker clustering method based on a generative model is proposed. It employs an incremental variant of variational Bayesian learning and provides probabilistic (non-deterministic) decisions for each input utterance, on the basis of the history of preceding utterances. It can be expected to be robust against errors in cluster estimation and the classification of utterances, and hence to be applicable to many real-time applications. Experimental results show that it produces 50% fewer classification errors than does a conventional online method. They also show that it is possible to reduce the number of speech recognition errors by combining the method with unsupervised speaker adaptation. Copyright © 2012 The Institute of Electronics, Information and Communication Engineers.

Cite

CITATION STYLE

APA

Koshinaka, T., Nagatomo, K., & Shinoda, K. (2012). Online speaker clustering using incremental learning of an ergodic hidden Markov model. IEICE Transactions on Information and Systems, E95-D(10), 2469–2478. https://doi.org/10.1587/transinf.E95.D.2469

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free