Improvement of the Fast Clustering Algorithm Improved by K -Means in the Big Data
Ting Xie, Ruihua Liu, Zhengyuan Wei
Source record
Source: Crossref
Published: Jan 1, 2020
DOI: 10.2478/amns.2020.1.00001
Open original source ↗Source abstract
Abstract Clustering as a fundamental unsupervised learning is considered an important method of data analysis, and K -means is demonstrably the most popular clustering algorithm. In this paper, we consider clustering on feature space to solve the low efficiency caused in the Big Data clustering by K -means. Different from the traditional methods, the algorithm guaranteed the consistency of the clustering accuracy before and after descending dimension, accelerated K -means when the clustering centeres and distance functions satisfy certain conditions, completely matched in the preprocessing step and clustering step, and improved the efficiency and accuracy. Experimental results have demonstrated the effectiveness of the proposed algorithm.
Evidence graph
No public relationships recorded yet.
Integrity note: This page is a factual metadata record created by deterministic ingestion. It is not a claim that the work moves a mathematical frontier or has been independently verified.