September 1, 2001

On feature distributional clustering for text categorization

Puntos clave

Los puntos clave no están disponibles para este artículo en este momento.

Resumen

We describe a text categorization approach that is based on a combination of feature distributional clusters with a support vector machine (SVM) classifier. Our feature selection approach employs distributional clustering of words via the recently introducedinformation bottleneck method, which generates a more efficientword-clusterrepresentation of documents. Combined with the classification power of an SVM, this method yields high performance text categorization that can outperform other recent methods in terms of categorization accuracy and representation efficiency. Comparing the accuracy of our method with other techniques, we observe significant dependency of the results on the data set. We discuss the potential reasons for this dependency.

Me gusta

Guardar

Cite This Study

Bekkerman et al. (Sat,) studied this question.

synapsesocial.com/papers/6a1d24291e7099f69104f2c5 https://doi.org/https://doi.org/10.1145/383952.383976

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

Me gusta

Guardar