p-values for classification

From MaRDI portal
\(p\)-values for classification



Abstract: Let (X,Y) be a random variable consisting of an observed feature vector XinmathcalX and an unobserved class label Yin1,2,...,L with unknown joint distribution. In addition, let mathcalD be a training data set consisting of n completely observed independent copies of (X,Y). Usual classification procedures provide point predictors (classifiers) widehatY(X,mathcalD) of Y or estimate the conditional distribution of Y given X. In order to quantify the certainty of classifying X we propose to construct for each heta=1,2,...,L a p-value piheta(X,mathcalD) for the null hypothesis that Y=heta, treating Y temporarily as a fixed parameter. In other words, the point predictor widehatY(X,mathcalD) is replaced with a prediction region for Y with a certain confidence. We argue that (i) this approach is advantageous over traditional approaches and (ii) any reasonable classifier can be modified to yield nonparametric p-values. We discuss issues such as optimality, single use and multiple use validity, as well as computational and graphical aspects.












This page was built for publication: \(p\)-values for classification

Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q1951759)