paper-with-me

홈 › Papers

Improved Bounds for Private and Robust Alignment

2025-12-29 · Wenqian Weng, Yi He, Xingyu Zhou arxiv

In this paper, we study the private and robust alignment of language models from a theoretical perspective by establishing upper bounds on the suboptimality gap in both offline and online settings. We consider preference labels subject to privacy constraints and/or adversarial corruption, and analyze two distinct interplays between them: privacy-first and corruption-first. For the privacy-only setting, we show that log loss with an MLE-style algorithm achieves near-optimal rates, in contrast to conventional wisdom. For the joint privacy-and-corruption setting, we first demonstrate that existing offline algorithms in fact provide stronger guarantees -- simultaneously in terms of corruption level and privacy parameters -- than previously known, which further yields improved bounds in the corruption-only regime. In addition, we also present the first set of results for private and robust online alignment. Our results are enabled by new uniform convergence guarantees for log loss and square loss under privacy and corruption, which we believe have broad applicability across learning theory and statistics.

📄 PDF Abstract BibTeX arXiv:2512.23816

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Closure Properties for Private Classification and Online Prediction

2020-03-10 · Noga Alon, Amos Beimel, Shay Moran, Uri Stemmer

Let~$\cH$ be a class of boolean functions and consider a {\it composed class} $\cH'$ that is derived from~$\cH$ using some arbitrary aggregation rule (for example, $\cH'$ may be the class of all 3-wise majority-votes of …

ClassificationGeneral ClassificationPAC learningPrediction

Improved Rates for Differentially Private Stochastic Convex Optimization with Heavy-Tailed Data

2021-06-02 · Gautam Kamath, Xingtu Liu, Huanyu Zhang

We study stochastic convex optimization with heavy-tailed data under the constraint of differential privacy (DP). Most prior work on this problem is restricted to the case where the loss function is Lipschitz. Instead, a…

Improved Accuracy for Private Continual Cardinality Estimation in Fully Dynamic Streams via Matrix Factorization

2026-01-05 · Joel Daniel Andersson, Palak Jain, Satchit Sivakumar arxiv

We study differentially-private statistics in the fully dynamic continual observation model, where many updates can arrive at each time step and updates to a stream can involve both insertions and deletions of an item. E…

Improved Bounds for Pure Private Agnostic Learning: Item-Level and User-Level Privacy

2024-07-30 · Bo Li, Wei Wang, Peng Ye

Machine Learning has made remarkable progress in a wide range of fields. In many scenarios, learning is performed on datasets involving sensitive information, in which privacy protection is essential for learning algorit…

Public-data Assisted Private Stochastic Optimization: Power and Limitations

2024-03-06 · Enayat Ullah, Michael Menart, Raef Bassily, Cristóbal Guzmán 외

We study the limits and capability of public-data assisted differentially private (PA-DP) algorithms. Specifically, we focus on the problem of stochastic convex optimization (SCO) with either labeled or unlabeled public …

Stochastic Optimization