Text processing for developing unrestricted tamil text to speech synthesis system

6Citations
Citations of this article
8Readers
Mendeley users who have this article in their library.

Abstract

In this Information and communication technology era, designing interactive computer systems that are effective, efficient, easy, and enjoyable to use is becoming increasingly important. Of the numerous ways explored by researchers to enhance Human-Computer Interaction, Text to Speech or Speech Synthesis affirms to be one such modality for developing better interfaces. The focal point here is to enhance the text processing module of Tamil speech synthesizer with an efficient and robust text normalizer and loan word identifier. Text normalization is performed on unrestricted Tamil text to convert non-standard words into standard words for the reduction of ambiguous utterances along the interim processing of the words. Loan words in Tamil text are identified in order to improve the pronunciation model of the Tamil speech synthesizer system. In this paper, we describe a 'semiotic classifier' based on decision list approach with which we are able to tackle many varieties of non-standard words. We also describe a 'loan/native word classifier' based on multiple linear regression which works efficiently even on shorter words of 3 syllables in length. In today's predominant Digital, Information-Communication Technology and Human-Computer Interaction era such profound text processors is imperative.

Cite

CITATION STYLE

APA

Rajendran, V., & Kumar, G. B. (2015). Text processing for developing unrestricted tamil text to speech synthesis system. Indian Journal of Science and Technology, 8(29). https://doi.org/10.17485/ijst/2015/v8i29/72294

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free