Randomized Nonlinear Component Analysis

Lopez-Paz, David; Sra, Suvrit; Smola, Alex; Ghahramani, Zoubin; Schölkopf, Bernhard

Statistics > Machine Learning

arXiv:1402.0119v1 (stat)

[Submitted on 1 Feb 2014 (this version), latest version 13 May 2014 (v2)]

Title:Randomized Nonlinear Component Analysis

Authors:David Lopez-Paz, Suvrit Sra, Alex Smola, Zoubin Ghahramani, Bernhard Schölkopf

View PDF

Abstract:Classical techniques such as Principal Component Analysis (PCA) and Canonical Correlation Analysis (CCA) are ubiquitous in statistics. However, these techniques only reveal linear relationships in data. Although nonlinear variants of PCA and CCA have been proposed, they are computationally prohibitive in the large scale.
In a separate strand of recent research, randomized methods have been proposed to construct features that help reveal nonlinear patterns in data. For basic tasks such as regression or classification, random features exhibit little or no loss in performance, while achieving dramatic savings in computational requirements.
In this paper we leverage randomness to design scalable new variants of nonlinear PCA and CCA; our ideas also extend to key multivariate analysis tools such as spectral clustering or LDA. We demonstrate our algorithms through experiments on real-world data, on which we compare against the state-of-the-art. Code in R implementing our methods is provided in the Appendix.

Subjects:	Machine Learning (stat.ML); Machine Learning (cs.LG)
Cite as:	arXiv:1402.0119 [stat.ML]
	(or arXiv:1402.0119v1 [stat.ML] for this version)
	https://doi.org/10.48550/arXiv.1402.0119

Submission history

From: David Lopez-Paz [view email]
[v1] Sat, 1 Feb 2014 19:54:06 UTC (312 KB)
[v2] Tue, 13 May 2014 16:41:11 UTC (530 KB)

Statistics > Machine Learning

Title:Randomized Nonlinear Component Analysis

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Statistics > Machine Learning

Title:Randomized Nonlinear Component Analysis

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators