PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
September 13, 2019ACM Computing Surveys243 citations

How Complex Is Your Classification Problem?

View Full Paper
ALAna Carolina LorenaLGLuís P. F. GarciaJLJens Lehmann

Key Points

Key points are not available for this paper at this time.

Abstract

Characteristics extracted from the training datasets of classification problems have proven to be effective predictors in a number of meta-analyses. Among them, measures of classification complexity can be used to estimate the difficulty in separating the data points into their expected classes. Descriptors of the spatial distribution of the data and estimates of the shape and size of the decision boundary are among the known measures for this characterization. This information can support the formulation of new data-driven pre-processing and pattern recognition techniques, which can in turn be focused on challenges highlighted by such characteristics of the problems. This article surveys and analyzes measures that can be extracted from the training datasets to characterize the complexity of the respective classification problems. Their use in recent literature is also reviewed and discussed, allowing to prospect opportunities for future work in the area. Finally, descriptions are given on an R package named Extended Complexity Library (ECoL) that implements a set of complexity measures and is made publicly available.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Lorena et al. (2019) studied this question.

synapsesocial.com/papers/69d84d955c3030ff03d19cd5https://doi.org/10.1145/3347711
Ask AI
Helpful
Bookmark
Share
View Full Paper