High-dimensional data clustering
From MaRDI portal
Abstract: Clustering in high-dimensional spaces is a difficult problem which is recurrent in many domains, for example in image analysis. The difficulty is due to the fact that high-dimensional data usually live in different low-dimensional subspaces hidden in the original space. This paper presents a family of Gaussian mixture models designed for high-dimensional data which combine the ideas of dimension reduction and parsimonious modeling. These models give rise to a clustering method based on the Expectation-Maximization algorithm which is called High-Dimensional Data Clustering (HDDC). In order to correctly fit the data, HDDC estimates the specific subspace and the intrinsic dimension of each group. Our experiments on artificial and real datasets show that HDDC outperforms existing methods for clustering high-dimensional data
Recommendations
- scientific article; zbMATH DE number 5280158
- Clustering high dimensional massive scientific datasets
- Clustering High-Dimensional Data via Feature Selection
- The challenges of clustering high dimensional data
- scientific article; zbMATH DE number 2112095
- Introduction to clustering large and high-dimensional data.
Cites work
- 10.1162/153244303322753616
- A classification EM algorithm for clustering and two stochastic versions
- A maximum likelihood methodology for clusterwise linear regression
- A mixture model for the classification of three-way proximity data
- A nonlinear PCA based on manifold approximation
- An Algorithm for Simultaneous Orthogonal Transformation of Several Positive Definite Symmetric Matrices to Nearly Diagonal Form
- ARPACK Users' Guide
- Automatische Klassifikation
- Detection and Characterization of Cluster Substructure I. Linear Structure: Fuzzy c-Lines
- Dimensionality reduction in quadratic discriminant analysis
- Discriminant Analysis with Singular Covariance Matrices: Methods and Applications to Spectroscopic Data
- Effect of dimensionality on discrimination
- Estimating Mixtures of Normal Distributions and Switching Regressions
- Estimating the dimension of a model
- Finite mixture models
- High-dimensional data clustering
- scientific article; zbMATH DE number 3126094 (Why is no real title available?)
- scientific article; zbMATH DE number 3567782 (Why is no real title available?)
- scientific article; zbMATH DE number 1059776 (Why is no real title available?)
- Model-Based Clustering, Discriminant Analysis, and Density Estimation
- Model-Based Gaussian and Non-Gaussian Clustering
- Modelling high-dimensional data by mixtures of factor analyzers
- On feature selection, curse-of-dimensionality and error probability in discriminant analysis
- Principal component analysis.
- Principal Curves
- Probabilistic models in cluster analysis
- The Discrimination Subspace Model
- Variable Selection for Model-Based Clustering
Cited in
(only showing first 100 items - show all)- Learning from partially supervised data using mixture models and belief functions
- A distance-relatedness dynamic model for clustering high dimensional data of arbitrary shapes and densities
- Editorial: Statistical learning methods including dimensionality reduction
- High-dimensional data clustering
- A hidden Markov model applied to the protein 3D structure analysis
- A mixture of generalized hyperbolic factor analyzers
- Model-based clustering of time series in group-specific functional subspaces
- Greedy clustering of count data through a mixture of multinomial PCA
- Simultaneous model-based clustering and visualization in the Fisher discriminative subspace
- A feature group weighting method for subspace clustering of high-dimensional data
- Model-based clustering of high-dimensional data: a review
- A hierarchical modeling approach for clustering probability density functions
- Model-based clustering for multivariate functional data
- Parsimonious skew mixture models for model-based clustering and classification
- Addressing overfitting and underfitting in Gaussian model-based clustering
- Location and scale mixtures of Gaussians with flexible tail behaviour: properties, inference and application to multivariate clustering
- Inverse regression approach to robust nonlinear high-to-low dimensional mapping
- Clustering of high values in random fields
- Variable selection methods for model-based clustering
- Mixtures of generalized hyperbolic distributions and mixtures of skew-t distributions for model-based clustering with incomplete data
- Clustering and classification via cluster-weighted factor analyzers
- Variational Bayes approximations for clustering via mixtures of normal inverse Gaussian distributions
- Functional data clustering by projection into latent generalized hyperbolic subspaces
- A Bayesian Fisher-EM algorithm for discriminative Gaussian subspace clustering
- Sparse mixture models inspired by ANOVA decompositions
- PCA reduced Gaussian mixture models with applications in superresolution
- Clustering and forecasting multiple functional time series
- Gaussian mixture model with an extended ultrametric covariance structure
- High-dimensional clustering via random projections
- A joint latent factor analyzer and functional subspace model for clustering multivariate functional data
- Mini-batch learning of exponential family finite mixture models
- Efficient mixture model for clustering of sparse high dimensional binary data
- Is-ClusterMPP: clustering algorithm through point processes and influence space towards high-dimensional data
- In the pursuit of sparseness: a new rank-preserving penalty for a finite mixture of factor analyzers
- Discriminative variable selection for clustering with the sparse Fisher-EM algorithm
- Robust supervised classification with mixture models: learning from data with uncertain labels
- Subspace clustering for the finite mixture of generalized hyperbolic distributions
- Large values of the clustering coefficient
- Clustering of imbalanced high-dimensional media data
- Functional data clustering: a survey
- Mixture model averaging for clustering
- Stable and visualizable Gaussian parsimonious clustering models
- The discriminative functional mixture model for a comparative analysis of bike sharing systems
- Model-based clustering
- Group-wise shrinkage estimation in penalized model-based clustering
- Dimensionally reduced model-based clustering through mixtures of factor mixture analyzers
- Adaptive mixture discriminant analysis for supervised learning with unobserved classes
- Sparse optimal discriminant clustering
- scientific article; zbMATH DE number 5817573 (Why is no real title available?)
- Application of affinity propagation for prototype sample detection, with application to face recognition
- Variable Selection for Clustering with Gaussian Mixture Models
- The generic subspace clustering model
- On clustering uncertain and structured data with Wasserstein barycenters and a geodesic criterion for the number of clusters
- Divisive clustering of high dimensional data streams
- High dimensional data clustering from a dynamical systems point of view
- HYPER-SPECTRAL DATA CLUSTERING METHOD BASED UPON THE SENSITIVE SUBSPACE
- scientific article; zbMATH DE number 5280158 (Why is no real title available?)
- Hierarchical subspace clustering
- Model-based clustering of longitudinal data
- Model-based clustering of high-dimensional data streams with online mixture of probabilistic PCA
- Clustering gene expression time course data using mixtures of multivariate \(t\)-distributions
- Analyzing state-dependent model-data comparison in multi-regime systems
- Theoretical and practical considerations on the convergence properties of the Fisher-EM algorithm
- Initializing the EM algorithm in Gaussian mixture models with an unknown number of components
- Statistical modeling of dissimilarity increments for \(d\)-dimensional data: application in partitional clustering
- scientific article; zbMATH DE number 1945788 (Why is no real title available?)
- Model-based classification via mixtures of multivariate \(t\)-distributions
- Model-based clustering, classification, and discriminant analysis of data with mixed type
- Robust discriminative clustering with sparse regularizers
- Clustering boundary pattern discovery for high dimensional space based on matrix model
- scientific article; zbMATH DE number 1832310 (Why is no real title available?)
- Latent simplex position model: high dimensional multi-view clustering with uncertainty quantification
- Heteroscedastic factor mixture analysis
- Holo-entropy based categorical data hierarchical clustering
- Compressive learning for patch-based image denoising
- Anderson relaxation test for intrinsic dimension selection in model-based clustering
- scientific article; zbMATH DE number 7578295 (Why is no real title available?)
- Optimal operator space pursuit: a framework for video sequence data analysis
- Reducing data dimension for cluster detection
- High-dimensional mixture models for unsupervised image denoising (HDMI)
- Projective clustering based on Parzen window technique
- Clustering high dimension, low sample size data using the maximal data piling distance
- Kernel discriminant analysis and clustering with parsimonious Gaussian process models
- A dual subspace parsimonious mixture of matrix normal distributions
- Cluster analysis with cellwise trimming and applications for the robust clustering of curves
- Tensor envelope mixture model for simultaneous clustering and multiway dimension reduction
- Clustering analysis of multivariate data: a weighted spatial ranks-based approach
- Image super-resolution with PCA reduced generalized Gaussian mixture models in materials science
- Parameter-wise co-clustering for high-dimensional data
- Functional data clustering via information maximization
- Frugal Gaussian clustering of huge imbalanced datasets through a bin-marginal approach
- Mixtures of modified t-factor analyzers for model-based clustering, classification, and discriminant analysis
- Dimensionally reduced mixtures of regression models
- Finite mixtures of matrix normal distributions for classifying three-way data
- A mixture of common skew-t factor analysers
- Flexible clustering of high-dimensional data via mixtures of joint generalized hyperbolic distributions
- Flexible mixture regression with the generalized hyperbolic distribution
- Parsimonious ultrametric Gaussian mixture models
- The parsimonious Gaussian mixture models with partitioned parameters and their application in clustering
- Model-based clustering with missing not at random data
This page was built for publication: High-dimensional data clustering
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q1020836)