Toward optimal probabilistic active learning using a Bayesian approach

From MaRDI portal



Abstract: Gathering labeled data to train well-performing machine learning models is one of the critical challenges in many applications. Active learning aims at reducing the labeling costs by an efficient and effective allocation of costly labeling resources. In this article, we propose a decision-theoretic selection strategy that (1) directly optimizes the gain in misclassification error, and (2) uses a Bayesian approach by introducing a conjugate prior distribution to determine the class posterior to deal with uncertainties. By reformulating existing selection strategies within our proposed model, we can explain which aspects are not covered in current state-of-the-art and why this leads to the superior performance of our approach. Extensive experiments on a large variety of datasets and different kernels validate our claims.


Active learning methods select the next input to be learned in order to make the learning more efficient under the supervised learning, that is, the learning of input-output relationships. This paper considers the active learning framework in a classification setting. Using the Dirichlet conjugate prior for the categorical model, the authors define the estimated risk difference and propose a new active learning strategy, expected probabilistic gain for active learning (xPAL). With some modifications, xPAL reduces to existing active learning methods, expected error reduction (EER), probabilistic active learning (PAL), and uncertainty sampling (US). The authors show these relationships theoretically and compare the performance of xPAL with the existing methods numerically. Although little is discussed on the connections to active learning frameworks based on statistical experimental design, xPAL provides a unifying framework for the recent active learning methods.




Cited in
(23)


Describes a project that uses

Uses Software






This page was built for publication: Toward optimal probabilistic active learning using a Bayesian approach

Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q2051315)