We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

cs.LG

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Computer Science > Machine Learning

Title: Debiased Contrastive Learning

Abstract: A prominent technique for self-supervised representation learning has been to contrast semantically similar and dissimilar pairs of samples. Without access to labels, dissimilar (negative) points are typically taken to be randomly sampled datapoints, implicitly accepting that these points may, in reality, actually have the same label. Perhaps unsurprisingly, we observe that sampling negative examples from truly different labels improves performance, in a synthetic setting where labels are available. Motivated by this observation, we develop a debiased contrastive objective that corrects for the sampling of same-label datapoints, even without knowledge of the true labels. Empirically, the proposed objective consistently outperforms the state-of-the-art for representation learning in vision, language, and reinforcement learning benchmarks. Theoretically, we establish generalization bounds for the downstream classification task.
Subjects: Machine Learning (cs.LG); Machine Learning (stat.ML)
Journal reference: Advances in Neural Information Processing Systems (2020)
Cite as: arXiv:2007.00224 [cs.LG]
  (or arXiv:2007.00224v3 [cs.LG] for this version)

Submission history

From: Ching-Yao Chuang [view email]
[v1] Wed, 1 Jul 2020 04:25:24 GMT (5024kb,D)
[v2] Sun, 5 Jul 2020 18:58:44 GMT (4953kb,D)
[v3] Wed, 21 Oct 2020 06:39:24 GMT (4901kb,D)

Link back to: arXiv, form interface, contact.