Distribution-Independent PAC Learning of Halfspaces with Massart Noise

Diakonikolas, Ilias; Gouleakis, Themis; Tzamos, Christos

Full-text links:

Download:

Current browse context:

cs.LG

< prev | next >

new | recent | 1906

Computer Science > Machine Learning

Title: Distribution-Independent PAC Learning of Halfspaces with Massart Noise

Authors: Ilias Diakonikolas, Themis Gouleakis, Christos Tzamos

(Submitted on 24 Jun 2019 (v1), last revised 10 Dec 2019 (this version, v2))

Abstract: We study the problem of {\em distribution-independent} PAC learning of halfspaces in the presence of Massart noise. Specifically, we are given a set of labeled examples $(\mathbf{x}, y)$ drawn from a distribution $\mathcal{D}$ on $\mathbb{R}^{d+1}$ such that the marginal distribution on the unlabeled points $\mathbf{x}$ is arbitrary and the labels $y$ are generated by an unknown halfspace corrupted with Massart noise at noise rate $\eta<1/2$. The goal is to find a hypothesis $h$ that minimizes the misclassification error $\mathbf{Pr}_{(\mathbf{x}, y) \sim \mathcal{D}} \left[ h(\mathbf{x}) \neq y \right]$.
We give a $\mathrm{poly}\left(d, 1/\epsilon \right)$ time algorithm for this problem with misclassification error $\eta+\epsilon$. We also provide evidence that improving on the error guarantee of our algorithm might be computationally hard. Prior to our work, no efficient weak (distribution-independent) learner was known in this model, even for the class of disjunctions. The existence of such an algorithm for halfspaces (or even disjunctions) has been posed as an open question in various works, starting with Sloan (1988), Cohen (1997), and was most recently highlighted in Avrim Blum's FOCS 2003 tutorial.

Subjects:	Machine Learning (cs.LG); Data Structures and Algorithms (cs.DS); Statistics Theory (math.ST); Machine Learning (stat.ML)
Cite as:	arXiv:1906.10075 [cs.LG]
	(or arXiv:1906.10075v2 [cs.LG] for this version)

Submission history

From: Ilias Diakonikolas [view email]
[v1] Mon, 24 Jun 2019 16:54:46 GMT (47kb,D)
[v2] Tue, 10 Dec 2019 16:46:00 GMT (48kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:1906.10075

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Machine Learning

Title: Distribution-Independent PAC Learning of Halfspaces with Massart Noise

Submission history