paper-with-me

홈 › Papers

PPI-SVRG: Unifying Prediction-Powered Inference and Variance Reduction for Semi-Supervised Optimization

2026-01-29 · Ruicheng Ao, Hongyu Chen, Haoyang Liu, David Simchi-Levi, Will Wei Sun arxiv

We study semi-supervised stochastic optimization when labeled data is scarce but predictions from pre-trained models are available. PPI and SVRG both reduce variance through control variates -- PPI uses predictions, SVRG uses reference gradients. We show they are mathematically equivalent and develop PPI-SVRG, which combines both. Our convergence bound decomposes into the standard SVRG rate plus an error floor from prediction uncertainty. The rate depends only on loss geometry; predictions affect only the neighborhood size. When predictions are perfect, we recover SVRG exactly. When predictions degrade, convergence remains stable but reaches a larger neighborhood. Experiments confirm the theory: PPI-SVRG reduces MSE by 43--52\% under label scarcity on mean estimation benchmarks and improves test accuracy by 2.7--2.9 percentage points on MNIST with only 10\% labeled data.

📄 PDF Abstract BibTeX arXiv:2601.21470

Code (0)

등록된 구현이 없습니다.

Tasks

Stochastic Optimization

Similar Papers 제목 키워드 기반

On Variance Reduction in Stochastic Gradient Descent and its Asynchronous Variants

2015-06-23 · NeurIPS 2015 12 · Sashank J. Reddi, Ahmed Hefny, Suvrit Sra, Barnabás Póczos 외

We study optimization algorithms based on variance reduction for stochastic gradient descent (SGD). Remarkable recent progress has been made in this direction through development of algorithms like SAG, SVRG, SAGA. These…

Compositional Stochastic Average Gradient for Machine Learning and Related Applications

2018-09-04 · Tsung-Yu Hsieh, Yasser EL-Manzalawy, Yiwei Sun, Vasant Honavar

Many machine learning, statistical inference, and portfolio optimization problems require minimization of a composition of expected value functions (CEVF). Of particular interest is the finite-sum versions of such compos…

BIG-bench Machine LearningPortfolio Optimization

Variance-reduced Zeroth-Order Methods for Fine-Tuning Language Models

2024-04-11 · Tanmay Gautam, Youngsuk Park, Hao Zhou, Parameswaran Raman 외

Fine-tuning language models (LMs) has demonstrated success in a wide array of downstream tasks. However, as LMs are scaled up, the memory requirements for backpropagation become prohibitively high. Zeroth-order (ZO) opti…

GPUIn-Context Learning

A Coefficient Makes SVRG Effective

2023-11-09 · Yida Yin, Zhiqiu Xu, Zhiyuan Li, Trevor Darrell 외

Stochastic Variance Reduced Gradient (SVRG), introduced by Johnson & Zhang (2013), is a theoretically compelling optimization method. However, as Defazio & Bottou (2019) highlight, its effectiveness in deep learning is y…

Deep Learningimage-classificationImage Classification

ASVRG: Accelerated Proximal SVRG

2018-10-07 · Fanhua Shang, Licheng Jiao, Kaiwen Zhou, James Cheng 외

This paper proposes an accelerated proximal stochastic variance reduced gradient (ASVRG) method, in which we design a simple and effective momentum acceleration trick. Unlike most existing accelerated stochastic variance…