K-means algorithm based Clustering for Big data
Abstract & Details
Research Area
Information Technology
Keywords
Clustering
Types of clustering
Classification
Data mining
big data.
Abstract
Clustering is a data mining technique used to place data elements into related groups without advance knowledge of the group definition. Clustering is a pro-cess of partitioning a set of data in a set of meaningful sub-classes, called cluster. In this paper, we propose to give a review of the most used clustering methods. First, we give an introduction about clustering methods, how they work and their main challenges. Second, we present the clustering methods with some comparisons including mainly the classical partitioning clustering methods like well-known k-means algorithms, Gaussian Mixture Modals and their variants, the classical hierarchical clustering methods. Clustering algorithms can be categorized into partition-based algorithms, hierarchical-based algorithms, density-based algorithms and grid-based algorithms. Partitioning clustering algorithm splits the data points into k partition, where each partition represents a cluster. Hierarchical clustering is a technique of clustering which divide the similar dataset by constructing a hierarchy of clusters. Density based algorithms and the cluster according to the regions which grow with high density. It is the one-scan algorithms. Grid Density based algorithm uses the multi resolution grid data structure and use dense grids to form clusters. Its main distinctiveness is the fastest processing time. In this survey paper, an analysis of clustering and its different techniques in data mining is done.
License
This work is licensed under a Creative
Commons
Attribution-ShareAlike 4.0 International License.
Author Information
| # | Name | Institute / Affiliation |
|---|---|---|
| 1 | Khevana Shah | L.D. College of Engineering |
How to Cite
Use the following formats to cite this article in your research.
APA Style
Shah, Khevana (2017). K-means algorithm based Clustering for Big data. International Journal of Advance Research and Innovative Ideas In Education, 3(5), 1156-1159.
MLA Style
Shah, Khevana. "K-means algorithm based Clustering for Big data." International Journal of Advance Research and Innovative Ideas In Education, vol. 3, no. 5, 2017, pp. 1156-1159.
IEEE Style
Khevana Shah, "K-means algorithm based Clustering for Big data," International Journal of Advance Research and Innovative Ideas In Education, vol. 3, no. 5, pp. 1156-1159, 2017.
Vancouver Style
Shah Khevana. K-means algorithm based Clustering for Big data. International Journal of Advance Research and Innovative Ideas In Education. 2017;3(5):1156-1159.
Harvard Style
Shah, Khevana (2017) 'K-means algorithm based Clustering for Big data', International Journal of Advance Research and Innovative Ideas In Education, 3(5), pp. 1156-1159.
Chicago Style
Shah, Khevana. "K-means algorithm based Clustering for Big data." International Journal of Advance Research and Innovative Ideas In Education 3, no. 5 (2017): 1156-1159.
Turabian Style
Shah, Khevana. "K-means algorithm based Clustering for Big data." International Journal of Advance Research and Innovative Ideas In Education 3, no. 5 (2017): 1156-1159.
Related Research
DIGITAL DIVIDE AND EQUITY IN ACCESS TO INTERNET: ITS IMPACT TO LEARNERS’ ACADEMIC ACHIEVEMENT
PDF Unavailable
A PHENOMENOLOGICAL STUDY ON THE CHALLENGES, AND COPING STRATEGIES OF SCHOOL HEADS IN USING TECHNOLOGY
PDF Unavailable
INFLUENCE OF TEACHER PERSONAL COMPETENCE AND SCHOOL LEADERSHIP ON STUDENT ACHIEVEMENT IN MEDIA AND INFORMATION LITERACY
PDF Unavailable
A Comprehensive Review of Blockchain in Automotive Data Tracking
PDF Unavailable
IoT-Based Elderly Emergency Health Monitoring System integrated with a Smart Ambulance mechanism
PDF Unavailable
Decentralized Voting System Using Ethereum Blockchain
PDF Unavailable