paper-with-me

홈 › Papers

Self-Training of Halfspaces with Generalization Guarantees under Massart Mislabeling Noise Model

2021-11-29 · Lies Hadjadj, Massih-Reza Amini, Sana Louhichi, Alexis Deschamps

We investigate the generalization properties of a self-training algorithm with halfspaces. The approach learns a list of halfspaces iteratively from labeled and unlabeled training data, in which each iteration consists of two steps: exploration and pruning. In the exploration phase, the halfspace is found sequentially by maximizing the unsigned-margin among unlabeled examples and then assigning pseudo-labels to those that have a distance higher than the current threshold. The pseudo-labeled examples are then added to the training set, and a new classifier is learned. This process is repeated until no more unlabeled examples remain for pseudo-labeling. In the pruning phase, pseudo-labeled samples that have a distance to the last halfspace greater than the associated unsigned-margin are then discarded. We prove that the misclassification error of the resulting sequence of classifiers is bounded and show that the resulting semi-supervised approach never degrades performance compared to the classifier learned using only the initial labeled training set. Experiments carried out on a variety of benchmarks demonstrate the efficiency of the proposed approach compared to state-of-the-art methods.

📄 PDF Abstract BibTeX arXiv:2111.14427

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Pruning 설명 없음

Similar Papers 제목 키워드 기반

Tight Generalization Bounds for Large-Margin Halfspaces

2025-02-19 · Kasper Green Larsen, Natascha Schalburg

We prove the first generalization bound for large-margin halfspaces that is asymptotically tight in the tradeoff between the margin, the fraction of training points with the given margin, the failure probability and the …

Generalization Bounds

Learning General Halfspaces with General Massart Noise under the Gaussian Distribution

2021-08-19 · Ilias Diakonikolas, Daniel M. Kane, Vasilis Kontonis, Christos Tzamos 외

We study the problem of PAC learning halfspaces on $\mathbb{R}^d$ with Massart noise under the Gaussian distribution. In the Massart model, an adversary is allowed to flip the label of each point $\mathbf{x}$ with unknow…

PAC learning

Classification Under Misspecification: Halfspaces, Generalized Linear Models, and Evolvability

2020-12-01 · NeurIPS 2020 12 · Sitan Chen, Frederic Koehler, Ankur Moitra, Morris Yau

In this paper, we revisit the problem of distribution-independently learning halfspaces under Massart noise with rate $\eta$. Recent work resolved a long-standing problem in this model of efficiently learning to error $\…

ClassificationFairnessGeneral ClassificationKnowledge Distillation

When are Local Queries Useful for Robust Learning?

2022-10-12 · Pascale Gourdeau, Varun Kanade, Marta Kwiatkowska, James Worrell

Distributional assumptions have been shown to be necessary for the robust learnability of concept classes when considering the exact-in-the-ball robust risk and access to random examples by Gourdeau et al. (2019). In thi…

Agnostic Learning of Halfspaces with Gradient Descent via Soft Margins

2020-10-01 · Spencer Frei, Yuan Cao, Quanquan Gu

We analyze the properties of gradient descent on convex surrogates for the zero-one loss for the agnostic learning of linear halfspaces. If $\mathsf{OPT}$ is the best classification error achieved by a halfspace, by appe…

General Classification