Decision tree instability and active learning

Kenneth Dwyer; Robert Holte

Conference ProceedingsOPEN ACCESS

Decision tree instability and active learning

Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) (2007) 4701 LNAI 128-139

DOI: 10.1007/978-3-540-74958-5_15

43Citations

61Readers

Abstract

Decision tree learning algorithms produce accurate models that can be interpreted by domain experts. However, these algorithms are known to be unstable - they can produce drastically different hypotheses from training sets that differ just slightly. This instability undermines the objective of extracting knowledge from the trees. In this paper, we study the instability of the C4.5 decision tree learner in the context of active learning. We introduce a new measure of decision tree stability, and define three aspects of active learning stability. Several existing active learning methods that use C4.5 as a component are compared empirically; it is determined that query-by-bagging yields trees that are more stable and accurate than those produced by competing methods. Also, an alternative splitting criterion, DKM, is found to improve the stability and accuracy of C4.5 in the active learning setting. © Springer-Verlag Berlin Heidelberg 2007.

Author supplied keywords

Cite

CITATION STYLE

APA

Dwyer, K., & Holte, R. (2007). Decision tree instability and active learning. In Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) (Vol. 4701 LNAI, pp. 128–139). Springer Verlag. https://doi.org/10.1007/978-3-540-74958-5_15

Decision tree instability and active learning

Abstract

Author supplied keywords

Cite

Register to see more suggestions