Ranking Median Regression: Learning to Order through Local Consensus

Clémençon, Stephan; Korba, Anna; Sibony, Eric

Full-text links:

Download:

Current browse context:

math

< prev | next >

new | recent | 1711

Mathematics > Statistics Theory

Title: Ranking Median Regression: Learning to Order through Local Consensus

Authors: Stephan Clémençon, Anna Korba, Eric Sibony

(Submitted on 31 Oct 2017 (v1), last revised 18 Dec 2017 (this version, v2))

Abstract: This article is devoted to the problem of predicting the value taken by a random permutation $\Sigma$, describing the preferences of an individual over a set of numbered items $\{1,\; \ldots,\; n\}$ say, based on the observation of an input/explanatory r.v. $X$ e.g. characteristics of the individual), when error is measured by the Kendall $\tau$ distance. In the probabilistic formulation of the 'Learning to Order' problem we propose, which extends the framework for statistical Kemeny ranking aggregation developped in \citet{CKS17}, this boils down to recovering conditional Kemeny medians of $\Sigma$ given $X$ from i.i.d. training examples $(X_1, \Sigma_1),\; \ldots,\; (X_N, \Sigma_N)$. For this reason, this statistical learning problem is referred to as \textit{ranking median regression} here. Our contribution is twofold. We first propose a probabilistic theory of ranking median regression: the set of optimal elements is characterized, the performance of empirical risk minimizers is investigated in this context and situations where fast learning rates can be achieved are also exhibited. Next we introduce the concept of local consensus/median, in order to derive efficient methods for ranking median regression. The major advantage of this local learning approach lies in its close connection with the widely studied Kemeny aggregation problem. From an algorithmic perspective, this permits to build predictive rules for ranking median regression by implementing efficient techniques for (approximate) Kemeny median computations at a local level in a tractable manner. In particular, versions of $k$-nearest neighbor and tree-based methods, tailored to ranking median regression, are investigated. Accuracy of piecewise constant ranking median regression rules is studied under a specific smoothness assumption for $\Sigma$'s conditional distribution given $X$.

Subjects:	Statistics Theory (math.ST); Machine Learning (stat.ML)
Cite as:	arXiv:1711.00070 [math.ST]
	(or arXiv:1711.00070v2 [math.ST] for this version)

Submission history

From: Anna Korba [view email]
[v1] Tue, 31 Oct 2017 19:40:40 GMT (389kb,D)
[v2] Mon, 18 Dec 2017 21:36:32 GMT (400kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> math > arXiv:1711.00070

Download:

Current browse context:

Change to browse by:

References & Citations

Bookmark

Mathematics > Statistics Theory

Title: Ranking Median Regression: Learning to Order through Local Consensus

Submission history