PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 26, 2022Artificial Intelligence Advances6 citationsOpen Access

Efficient Parallel Processing of k-Nearest Neighbor Queries by Using a Centroid-based and Hierarchical Clustering Algorithm

View Full Paper
EGElaheh GavagsazIslamic Azad University South Tehran Branch

Key Points

Key points are not available for this paper at this time.

Abstract

The k-Nearest Neighbor method is one of the most popular techniques for both classification and regression purposes. Because of its operation, the application of this classification may be limited to problems with a certain number of instances, particularly, when run time is a consideration. However, the classification of large amounts of data has become a fundamental task in many real-world applications. It is logical to scale the k-Nearest Neighbor method to large scale datasets. This paper proposes a new k-Nearest Neighbor classification method (KNN-CCL) which uses a parallel centroid-based and hierarchical clustering algorithm to separate the sample of training dataset into multiple parts. The introduced clustering algorithm uses four stages of successive refinements and generates high quality clusters. The k-Nearest Neighbor approach subsequently makes use of them to predict the test datasets. Finally, sets of experiments are conducted on the UCI datasets. The experimental results confirm that the proposed k-Nearest Neighbor classification method performs well with regard to classification accuracy and performance.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Elaheh Gavagsaz (2022) studied this question.

synapsesocial.com/papers/6a129df48793652519a641d8https://doi.org/10.30564/aia.v4i1.4668
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Efficient parallel kNN joins for large data in MapReduce2012 · 253 citations
  2. 2Similarity-based Classification: Concepts and Algorithms2009 · 294 citations
  3. 3An intermediate data placement algorithm for load balancing in Spark computing environment2016 · 75 citations
  4. 4Nearest neighbor pattern classification1967 · 16,720 citations
  5. 5Density-Based Clustering2017 · 11 citations