Capturing patterns of variation unique to a specific dataset

Tu, Robin; Foss, Alexander H.; Zhao, Sihai D.

Full-text links:

Download:

Current browse context:

cs.LG

< prev | next >

new | recent | 2104

Computer Science > Machine Learning

Title: Capturing patterns of variation unique to a specific dataset

Authors: Robin Tu, Alexander H. Foss, Sihai D. Zhao

(Submitted on 16 Apr 2021)

Abstract: Capturing patterns of variation present in a dataset is important in exploratory data analysis and unsupervised learning. Contrastive dimension reduction methods, such as contrastive principal component analysis (cPCA), find patterns unique to a target dataset of interest by contrasting with a carefully chosen background dataset representing unwanted or uninteresting variation. However, such methods typically require a tuning parameter that governs the level of contrast, and it is unclear how to choose this parameter objectively. Furthermore, it is frequently of interest to contrast against multiple backgrounds, which is difficult to accomplish with existing methods. We propose unique component analysis (UCA), a tuning-free method that identifies low-dimensional representations of a target dataset relative to one or more comparison datasets. It is computationally efficient even with large numbers of features. We show in several experiments that UCA with a single background dataset achieves similar results compared to cPCA with various tuning parameters, and that UCA with multiple individual background datasets is superior to both cPCA with any single background data and cPCA with a pooled background dataset.

Subjects:	Machine Learning (cs.LG); Methodology (stat.ME)
Cite as:	arXiv:2104.08157 [cs.LG]
	(or arXiv:2104.08157v1 [cs.LG] for this version)

Submission history

From: Robin Tu [view email]
[v1] Fri, 16 Apr 2021 15:07:32 GMT (7192kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2104.08157

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Machine Learning

Title: Capturing patterns of variation unique to a specific dataset

Submission history