With the appearance of Unicode encoding, content in Indian dialects is continually expanding on the internet. Gathering of artistic writings in Punjabi language, particularly poetry, is expanding day by day on the web. In this way, grouping of poems, as indicated by topic, is a critical errand. Classification of poems is very challenging in computational linguistic point of view. Manual collection of 240 poems in four categories is done and passed to pre-processing phase. Gain ratio is used for ranking features. K-nearest neighbour (k-KNN), Naïve Bayes (NB), support vector machine (SVM) and hyperpipes (HP) are trained and tested. Outcomes indicate that Naïve Bayes outperformed all other classifiers utilising 60% top ranked features and hyperpipes is the least efficient classifier. Result additionally demonstrates 15% increase in accuracy by utilising gain ratio as feature selection technique.
No takes yet. Share an insight, caveat, or question.
Kaur et al. (2016) studied this question.
Synapse has enriched 2 closely related papers on similar clinical questions. Consider them for comparative context: