K-means clustering to improve the accuracy of Decision tree response classification

29Citations
Citations of this article
22Readers
Mendeley users who have this article in their library.

Abstract

The use of deep generation with statistical-based surface generation merits from response utterances readily available from corpus. Representation and quality of the instance data are the foremost factors that affect classification accuracy of the statistical-based method. Thus, in classification task, any irrelevant or unreliable tagging of response classes represented will result in low accuracy. This study focused on improving dialogue act classification of a user utterance into a response class by clustering the semantic and pragmatic features extracted from each user utterance. A Decision tree approach is used to classify 64 mixed-initiative, transaction dialogue corpus in theater domain. The experiment shows that by using clustering technique in preprocessing stage for re-tagging response classes, the Decision tree is able to achieve 97.5% recognition accuracy in classification, better than the 81.95% recognition accuracy when using Decision tree alone. © 2009 Asian Network for Scientific Information.

Cite

CITATION STYLE

APA

Ali, S. A., Sulaiman, N., Mustapha, A., & Mustapha, N. (2009). K-means clustering to improve the accuracy of Decision tree response classification. Information Technology Journal, 8(8), 1256–1262. https://doi.org/10.3923/itj.2009.1256.1262

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free