Variable importance in binary regression trees and forests

360Citations
Citations of this article
250Readers
Mendeley users who have this article in their library.

Abstract

We characterize and study variable importance (VIMP) and pairwise variable associations in binary regression trees. A key component involves the node mean squared error for a quantity we refer to as a maximal subtree. The theory naturally extends from single trees to ensembles of trees and applies to methods like random forests. This is useful because while importance values from random forests are used to screen variables, for example they are used to filter high throughput genomic data in Bioinformatics, very little theory exists about their properties. © 2007, Ashdin Publishing. All rights reserved.

Author supplied keywords

Cite

CITATION STYLE

APA

Ishwaran, H. (2007). Variable importance in binary regression trees and forests. Electronic Journal of Statistics, 1, 519–537. https://doi.org/10.1214/07-EJS039

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free