November 2018
Beginner to intermediate
568 pages
15h 59m
English
Clustering is an unsupervised data science technique where the records in a dataset are organized into different logical groupings. The data are grouped in such a way that records inside the same group are more similar than records outside the group. Clustering has a wide variety of applications ranging from market segmentation to customer segmentation, electoral grouping, web analytics, and outlier detection. Clustering is also used as a data compression technique and data preprocessing technique for supervised tasks. Many different data science approaches are available to cluster the data and are developed based on proximity between the records, density in the dataset, or novel application of neural networks. k
Read now
Unlock full access