Search CORE

38 research outputs found

Efficient Learning of Linear Separators under Bounded Noise

Author: Awasthi Pranjal
Balcan Maria-Florina
Haghtalab Nika
Urner Ruth
Publication venue
Publication date: 12/03/2015
Field of study

We study the learnability of linear separators in

\Re^d

in the presence of bounded (a.k.a Massart) noise. This is a realistic generalization of the random classification noise model, where the adversary can flip each example

x

with probability

\eta(x) \leq \eta

. We provide the first polynomial time algorithm that can learn linear separators to arbitrarily small excess error in this noise model under the uniform distribution over the unit ball in

\Re^d

, for some constant value of

\eta

. While widely studied in the statistical learning theory community in the context of getting faster convergence rates, computationally efficient algorithms in this model had remained elusive. Our work provides the first evidence that one can indeed design algorithms achieving arbitrarily small excess error in polynomial time under this realistic noise model and thus opens up a new and exciting line of research. We additionally provide lower bounds showing that popular algorithms such as hinge loss minimization and averaging cannot lead to arbitrarily small excess error under Massart noise, even under the uniform distribution. Our work instead, makes use of a margin based technique developed in the context of active learning. As a result, our algorithm is also an active learning algorithm with label complexity that is only a logarithmic the desired excess error

\epsilon

arXiv.org e-Print Archive

MPG.PuRe

A PTAS for Agnostically Learning Halfspaces

Author: Daniely Amit
Publication venue
Publication date: 01/01/2015
Field of study

We present a PTAS for agnostically learning halfspaces w.r.t. the uniform distribution on the

d

dimensional sphere. Namely, we show that for every

\mu>0

there is an algorithm that runs in time

\mathrm{poly}(d,\frac{1}{\epsilon})

, and is guaranteed to return a classifier with error at most

(1+\mu)\mathrm{opt}+\epsilon

, where

\mathrm{opt}

is the error of the best halfspace classifier. This improves on Awasthi, Balcan and Long [ABL14] who showed an algorithm with an (unspecified) constant approximation ratio. Our algorithm combines the classical technique of polynomial regression (e.g. [LMN89, KKMS05]), together with the new localization technique of [ABL14]

arXiv.org e-Print Archive

CiteSeerX

The Power of Localization for Efficiently Learning Linear Separators with Noise

Author: Awasthi Pranjal
Balcan Maria Florina
Long Philip M.
Publication venue
Publication date: 03/06/2018
Field of study

We introduce a new approach for designing computationally efficient learning algorithms that are tolerant to noise, and demonstrate its effectiveness by designing algorithms with improved noise tolerance guarantees for learning linear separators. We consider both the malicious noise model and the adversarial label noise model. For malicious noise, where the adversary can corrupt both the label and the features, we provide a polynomial-time algorithm for learning linear separators in

\Re^d

under isotropic log-concave distributions that can tolerate a nearly information-theoretically optimal noise rate of

\eta = \Omega(\epsilon)

. For the adversarial label noise model, where the distribution over the feature vectors is unchanged, and the overall probability of a noisy label is constrained to be at most

\eta

, we also give a polynomial-time algorithm for learning linear separators in

\Re^d

under isotropic log-concave distributions that can handle a noise rate of

\eta = \Omega\left(\epsilon\right)

. We show that, in the active learning model, our algorithms achieve a label complexity whose dependence on the error parameter

\epsilon

is polylogarithmic. This provides the first polynomial-time active learning algorithm for learning linear separators in the presence of malicious noise or adversarial label noise.Comment: Contains improved label complexity analysis communicated to us by Steve Hannek

arXiv.org e-Print Archive

FigShare