A survey on feature weighting based K-means algorithms

From MaRDI portal
Publication:333337

DOI10.1007/S00357-016-9208-4zbMATH Open1349.62291arXiv1601.03483OpenAlexW2238173180MaRDI QIDQ333337FDOQ333337

Renato Cordeiro de Amorim

Publication date: 28 October 2016

Published in: Journal of Classification (Search for Journal in Brave)

Abstract: In a real-world data set there is always the possibility, rather high in our opinion, that different features may have different degrees of relevance. Most machine learning algorithms deal with this fact by either selecting or deselecting features in the data preprocessing phase. However, we maintain that even among relevant features there may be different degrees of relevance, and this should be taken into account during the clustering process. With over 50 years of history, K-Means is arguably the most popular partitional clustering algorithm there is. The first K-Means based clustering algorithm to compute feature weights was designed just over 30 years ago. Various such algorithms have been designed since but there has not been, to our knowledge, a survey integrating empirical evidence of cluster recovery ability, common flaws, and possible directions for future research. This paper elaborates on the concept of feature weighting and addresses these issues by critically analysing some of the most popular, or innovative, feature weighting mechanisms based in K-Means.


Full work available at URL: https://arxiv.org/abs/1601.03483




Recommendations




Cites Work


Cited In (14)

Uses Software





This page was built for publication: A survey on feature weighting based K-means algorithms

Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q333337)