Improving the precision of classification trees
From MaRDI portal
Abstract: Besides serving as prediction models, classification trees are useful for finding important predictor variables and identifying interesting subgroups in the data. These functions can be compromised by weak split selection algorithms that have variable selection biases or that fail to search beyond local main effects at each node of the tree. The resulting models may include many irrelevant variables or select too few of the important ones. Either eventuality can lead to erroneous conclusions. Four techniques to improve the precision of the models are proposed and their effectiveness compared with that of other algorithms, including tree ensembles, on real and simulated data sets.
Recommendations
Cites work
- 10.1162/153244304322972694
- A comparison of prediction accuracy, complexity, and training time of thirty-three old and new classification algorithms
- An unbiased method for constructing multilabel classification trees
- Functional trees
- FUNDAMENTAL GROUPS
- Heuristics of instability and stabilization in model selection
- scientific article; zbMATH DE number 3860199 (Why is no real title available?)
- scientific article; zbMATH DE number 47310 (Why is no real title available?)
- scientific article; zbMATH DE number 1089157 (Why is no real title available?)
- scientific article; zbMATH DE number 1149421 (Why is no real title available?)
- scientific article; zbMATH DE number 3434963 (Why is no real title available?)
- scientific article; zbMATH DE number 1547349 (Why is no real title available?)
- scientific article; zbMATH DE number 1779487 (Why is no real title available?)
- scientific article; zbMATH DE number 2200689 (Why is no real title available?)
- Problems in the Analysis of Survey Data, and a Proposal
- Random forests
- The Distribution of Chi-Square
- Tree-Structured Classification Via Generalized Discriminant Analysis
- Unbiased split selection for classification trees based on the Gini index
- Unbiased variable selection for classification trees with multivariate responses
- Using \(k\)-nearest-neighbor classification in the leaves of a tree
Cited in
(51)- Classification tree analysis using TARGET
- Calibration and refinement for classification trees
- Nonparametric variable selection and classification: the CATCH algorithm
- Accurate ensemble pruning with PL-bagging
- An alternative pruning based approach to unbiased recursive partitioning
- Modeling threshold interaction effects through the logistic classification trunk
- A procedure for improving generalization in classification trees
- PPtree: projection pursuit classification tree
- Regression trees for longitudinal and multiresponse data
- The reliability of classification of terminal nodes in GUIDE decision tree to predict the nonalcoholic fatty liver disease
- An omnibus test for detection of subgroup treatment effects via data partitioning
- Subgroup identification with classification and regression tree-based algorithms: an application to the ball state adult fitness longitudinal study
- Training trees on tails with applications to portfolio choice
- Count regression trees
- Double random forest
- A new classification tree method with interaction detection capability
- Canonical forest
- Predicting missing values: a comparative study on non-parametric approaches for imputation
- T3C: improving a decision tree classification algorithm's interval splits on continuous attributes
- Model trees with topic model preprocessing: an approach for data journalism illustrated with the WikiLeaks Afghanistan war logs
- scientific article; zbMATH DE number 1629814 (Why is no real title available?)
- Unified Noncrossing Multiple Quantile Regressions Tree
- Supervised classification and mathematical optimization
- scientific article; zbMATH DE number 1089157 (Why is no real title available?)
- scientific article; zbMATH DE number 1983958 (Why is no real title available?)
- Appraisal of performance of three tree-based classification methods
- scientific article; zbMATH DE number 1881089 (Why is no real title available?)
- Composite large margin classifiers with latent subclasses for heterogeneous biomedical data
- Seemingly unrelated regression tree
- bsnsing: A Decision Tree Induction Method Based on Recursive Optimal Boolean Rule Composition
- Disjunctive Rule Lists
- Phylogenetic tree selection by the adjustedk-means approach
- A note on the interpretation of tree-based regression models
- Gaining insight with recursive partitioning of generalized linear models
- Inductive Logic Programming
- Node harvest
- Model-based recursive partitioning algorithm to penalized non-crossing multiple quantile regression for the right-censored data
- Binary multi-layer classifier
- On hybrid tree-based methods for short-term insurance claims
- A nonparametric method for classification trees using grouped covariates
- New optimization models for optimal classification trees
- Partially Bayesian variable selection in classification trees
- A platform for comparing subgroup identification methodologies
- Model-based recursive partitioning for subgroup analyses
- Random forest variable importance-based selection algorithm in class imbalance problem
- A regression tree approach to identifying subgroups with differential treatment effects
- Identification of subgroups with differential treatment effects for longitudinal and multiresponse variables
- Accounting for shared covariates in semiparametric Bayesian additive regression trees
- Maximizing adjusted covariance: new supervised dimension reduction for classification
- A weight-adjusted voting algorithm for ensembles of classifiers
- A note on split selection bias in classification trees
This page was built for publication: Improving the precision of classification trees
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q965140)