paper-with-me

홈 › Papers

DPack: Efficiency-Oriented Privacy Budget Scheduling

2022-12-26 · Pierre Tholoniat, Kelly Kostopoulou, Mosharaf Chowdhury, Asaf Cidon, Roxana Geambasu, Mathias Lécuyer, Junfeng Yang

Machine learning (ML) models can leak information about users, and differential privacy (DP) provides a rigorous way to bound that leakage under a given budget. This DP budget can be regarded as a new type of compute resource in workloads of multiple ML models training on user data. Once it is used, the DP budget is forever consumed. Therefore, it is crucial to allocate it most efficiently to train as many models as possible. This paper presents the scheduler for privacy that optimizes for efficiency. We formulate privacy scheduling as a new type of multidimensional knapsack problem, called privacy knapsack, which maximizes DP budget efficiency. We show that privacy knapsack is NP-hard, hence practical algorithms are necessarily approximate. We develop an approximation algorithm for privacy knapsack, DPack, and evaluate it on microbenchmarks and on a new, synthetic private-ML workload we developed from the Alibaba ML cluster trace. We show that DPack: (1) often approaches the efficiency-optimal schedule, (2) consistently schedules more tasks compared to a state-of-the-art privacy scheduling algorithm that focused on fairness (1.3-1.7x in Alibaba, 1.0-2.6x in microbenchmarks), but (3) sacrifices some level of fairness for efficiency. Therefore, using DPack, DP ML operators should be able to train more models on the same amount of user data while offering the same privacy guarantee to their users.

📄 PDF Abstract BibTeX arXiv:2212.13228

Code (2)

columbia/alibaba-dp-workload 공식 구현
columbia/dpack 공식 구현

Tasks

FairnessScheduling

Similar Papers 제목 키워드 기반

DPBalance: Efficient and Fair Privacy Budget Scheduling for Federated Learning as a Service

2024-02-15 · Yu Liu, Zibo Wang, Yifei Zhu, Chen Chen

Federated learning (FL) has emerged as a prevalent distributed machine learning scheme that enables collaborative model training without aggregating raw data. Cloud service providers further embrace Federated Learning as…

FairnessFederated LearningScheduling

Budgeted Policy Learning for Task-Oriented Dialogue Systems

2019-06-02 · ACL 2019 7 · Zhirui Zhang, Xiujun Li, Jianfeng Gao, Enhong Chen

This paper presents a new approach that extends Deep Dyna-Q (DDQ) by incorporating a Budget-Conscious Scheduling (BCS) to best utilize a fixed, small amount of user interactions (budget) for learning task-oriented dialog…

SchedulingTask-Oriented Dialogue Systems

Over-the-Air Federated Averaging with Limited Power and Privacy Budgets

2023-05-05 · Na Yan, Kezhi Wang, Cunhua Pan, Kok Keong Chai 외

To jointly overcome the communication bottleneck and privacy leakage of wireless federated learning (FL), this paper studies a differentially private over-the-air federated averaging (DP-OTA-FedAvg) system with a limited…

Federated LearningScheduling

WorldPack: Dynamic Frame Compression for Long-context Video World Modeling

2025-12-02 · Yuta Oshima, Yusuke Iwasawa, Masahiro Suzuki, Yutaka Matsuo 외 arxiv

Video world models have attracted significant attention for their ability to produce high-fidelity future visual observations conditioned on past observations and navigation actions. However, achieving temporally and spa…

Spatial Reasoning

Device Scheduling with Fast Convergence for Wireless Federated Learning

2019-11-03 · Wenqi Shi, Sheng Zhou, Zhisheng Niu

Owing to the increasing need for massive data analysis and model training at the network edge, as well as the rising concerns about the data privacy, a new distributed training framework called federated learning (FL) ha…

Federated LearningScheduling