PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 30, 2016Machine Learning and Applications An International Journal368 citationsOpen Access

A Survey on Similarity Measures in Text Mining

VMVijaymeena M.KKKK. Kavitha

Key Points

Key points are not available for this paper at this time.

Abstract

The Volume of text resources have been increasing in digital libraries and internet. Organizing these text documents has become a practical need. For organizing great number of objects into small or minimum number of coherent groups automatically, Clustering technique is used. These documents are widely used for information retrieval and Natural Language processing tasks. Different Clustering algorithms require a metric for quantifying how dissimilar two given documents are. This difference is often measured by similarity measure such as Euclidean distance, Cosine similarity etc. The similarity measure process in text mining can be used to identify the suitable clustering algorithm for a specific problem. This survey discusses the existing works on text similarity by partitioning them into three significant approaches; String-based, Knowledge based and Corpus-based similarities.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

M.K et al. (2016) studied this question.

synapsesocial.com/papers/69f8a94ea1b2dcd77ea63622https://doi.org/10.5121/mlaij.2016.3103
Ask AI
Helpful
Bookmark
Share
View Full Paper