Abstract
Decision tree-based clustering and parameter estimation are essential steps in the training part of an HMM-based speech synthesis system. These two steps are usually performed based on the maximum likelihood (ML) criterion. However, one of the drawbacks of the ML criterion is that it is sensitive to outliers which usually result in quality degradation of the synthesized speech. In this letter, we propose an approach to detect and remove outliers for HMM-based speech synthesis. Experimental results show that the proposed approach can improve the synthetic speech, particularly when the available training speech database is insufficient. Copyright © 2012 The Institute of Electronics, Information and Communication Engineers.
Author supplied keywords
Cite
CITATION STYLE
Hong, D. H., Sung, J. S., Oh, K. H., & Kim, N. S. (2012). Outlier detection and removal for HMM-based speech synthesis with an insufficient speech database. IEICE Transactions on Information and Systems, E95-D(9), 2351–2354. https://doi.org/10.1587/transinf.E95.D.2351
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.