PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
November 23, 2002175 citations

Dimensionality reduction of unsupervised data

View Full Paper
MDManoranjan DashNorthwestern UniversityHLHua LiuSinopec (China)JYJia YaoNingbo University

Key Points

Key points are not available for this paper at this time.

Abstract

Dimensionality reduction is an important problem for efficient handling of large databases. Many feature selection methods exist for supervised data having class information. Little work has been done for dimensionality reduction of unsupervised data in which class information is not available. Principal component analysis (PCA) is often used. However, PCA creates new features. It is difficult to obtain intuitive understanding of the data using the new features only. We are concerned with the problem of determining and choosing the important original features for unsupervised data. Our method is based on the observation that removing an irrelevant feature from the feature set may not change the underlying concept of the data, but not so otherwise. We propose an entropy measure for ranking features, and conduct extensive experiments to show that our method is able to find the important features. Also it compares well with a similar feature ranking method (Relief) that requires class information unlike our method.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Dash et al. (2002) studied this question.

synapsesocial.com/papers/6a2417154e11633d95ab339ehttps://doi.org/10.1109/tai.1997.632300
Ask AI
Helpful
Bookmark
Share
View Full Paper