A novel graph-based language-independent stemming algorithm suitable for information retrieval is proposed in this article. The main features of the algorithm are retrieval effectiveness, generality, and computational efficiency. We test our approach on seven languages (using collections from the TREC, CLEF, and FIRE evaluation platforms) of varying morphological complexity. Significant performance improvement over plain word-based retrieval, three other language-independent morphological normalizers, as well as rule-based stemmers is demonstrated.
No takes yet. Share an insight, caveat, or question.
Paik et al. (2011) studied this question.
Synapse has enriched 4 closely related papers on similar clinical questions. Consider them for comparative context: