We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:


Current browse context:


Change to browse by:

References & Citations


(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo ScienceWISE logo

Statistics > Machine Learning

Title: Sketching Datasets for Large-Scale Learning (long version)

Abstract: This article considers "compressive learning," an approach to large-scale machine learning where datasets are massively compressed before learning (e.g., clustering, classification, or regression) is performed. In particular, a "sketch" is first constructed by computing carefully chosen nonlinear random features (e.g., random Fourier features) and averaging them over the whole dataset. Parameters are then learned from the sketch, without access to the original dataset. This article surveys the current state-of-the-art in compressive learning, including the main concepts and algorithms, their connections with established signal-processing methods, existing theoretical guarantees -- on both information preservation and privacy preservation, and important open problems.
Subjects: Machine Learning (stat.ML); Information Theory (cs.IT); Machine Learning (cs.LG)
Cite as: arXiv:2008.01839 [stat.ML]
  (or arXiv:2008.01839v3 [stat.ML] for this version)

Submission history

From: Philip Schniter [view email]
[v1] Tue, 4 Aug 2020 21:29:05 GMT (17703kb,D)
[v2] Tue, 19 Jan 2021 20:41:41 GMT (20095kb,D)
[v3] Thu, 24 Jun 2021 21:36:36 GMT (19141kb,D)

Link back to: arXiv, form interface, contact.