We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

stat.ML

Change to browse by:

References & Citations

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Statistics > Machine Learning

Title: Kernel k-Groups via Hartigan's Method

Abstract: Energy statistics was proposed by Sz\' ekely in the 80's inspired by Newton's gravitational potential in classical mechanics and it provides a model-free hypothesis test for equality of distributions. In its original form, energy statistics was formulated in Euclidean spaces. More recently, it was generalized to metric spaces of negative type. In this paper, we consider a formulation for the clustering problem using a weighted version of energy statistics in spaces of negative type. We show that this approach leads to a quadratically constrained quadratic program in the associated kernel space, establishing connections with graph partitioning problems and kernel methods in machine learning. To find local solutions of such an optimization problem, we propose kernel k-groups, which is an extension of Hartigan's method to kernel spaces. Kernel k-groups is cheaper than spectral clustering and has the same computational cost as kernel k-means (which is based on Lloyd's heuristic) but our numerical results show an improved performance, especially in higher dimensions. Moreover, we verify the efficiency of kernel k-groups in community detection in sparse stochastic block models which has fascinating applications in several areas of science.
Comments: several improvements; connections with community detection and stochastic block model. Matches published version
Subjects: Machine Learning (stat.ML); Computer Vision and Pattern Recognition (cs.CV); Data Structures and Algorithms (cs.DS); Machine Learning (cs.LG); Statistics Theory (math.ST)
Journal reference: IEEE Transactions on Pattern Analysis and Machine Intelligence, 2020
DOI: 10.1109/TPAMI.2020.2998120
Cite as: arXiv:1710.09859 [stat.ML]
  (or arXiv:1710.09859v4 [stat.ML] for this version)

Submission history

From: Guilherme França [view email]
[v1] Thu, 26 Oct 2017 18:38:28 GMT (2323kb,D)
[v2] Tue, 14 Aug 2018 14:02:55 GMT (1871kb,D)
[v3] Mon, 9 Dec 2019 15:29:58 GMT (2437kb,D)
[v4] Thu, 11 Jun 2020 19:57:09 GMT (2431kb,D)

Link back to: arXiv, form interface, contact.