PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
June 14, 2009826 citationsOpen Access

Multi-view clustering via canonical correlation analysis

KCKamalika ChaudhuriSKSham M. KakadeKLKaren Livescu

Key Points

Key points are not available for this paper at this time.

Abstract

Clustering data in high-dimensions is believed to be a hard problem in general. A number of efficient clustering algorithms developed in recent years address this problem by projecting the data into a lower-dimensional subspace, e. g. via Principal Components Analysis (PCA) or random projections, before clustering. Such techniques typically require stringent requirements on the separation between the cluster means (in order for the algorithm to be be successful). , we show how using multiple views of the data can relax these stringent requirements. We use Canonical Correlation Analysis (CCA) to project the data in each view to a lower-dimensional subspace. Under the assumption that conditioned on the cluster label the views are uncorrelated, we show that the separation conditions required for the algorithm to be successful are rather mild (significantly weaker than those of prior results in the literature). We provide results for mixture of Gaussians, mixtures of log concave distributions, and mixtures of product distributions.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Chaudhuri et al. (2009) studied this question.

synapsesocial.com/papers/69d995318d5c517421a3c030https://doi.org/10.1145/1553374.1553391
Ask AI
Helpful
Bookmark
Share
View Full Paper