We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

cs.LG

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Computer Science > Machine Learning

Title: Optimal 1-NN Prototypes for Pathological Geometries

Abstract: Using prototype methods to reduce the size of training datasets can drastically reduce the computational cost of classification with instance-based learning algorithms like the k-Nearest Neighbour classifier. The number and distribution of prototypes required for the classifier to match its original performance is intimately related to the geometry of the training data. As a result, it is often difficult to find the optimal prototypes for a given dataset, and heuristic algorithms are used instead. However, we consider a particularly challenging setting where commonly used heuristic algorithms fail to find suitable prototypes and show that the optimal prototypes can instead be found analytically. We also propose an algorithm for finding nearly-optimal prototypes in this setting, and use it to empirically validate the theoretical results.
Comments: 8 pages
Subjects: Machine Learning (cs.LG); Machine Learning (stat.ML)
DOI: 10.7717/peerj-cs.464
Cite as: arXiv:2011.00228 [cs.LG]
  (or arXiv:2011.00228v1 [cs.LG] for this version)

Submission history

From: Ilia Sucholutsky [view email]
[v1] Sat, 31 Oct 2020 10:15:08 GMT (190kb,D)

Link back to: arXiv, form interface, contact.