paper-with-me

Papers

Adaptive Sampling for Deep Learning via Efficient Nonparametric Proxies

2023-11-22 · Shabnam Daghaghi, Benjamin Coleman, Benito Geordie, Anshumali Shrivastava

Data sampling is an effective method to improve the training speed of neural networks, with recent results demonstrating that it can even break the neural scaling laws. These results critically rely on high-quality scores to estimate the importance of an input to the network. We observe that there are two dominant strategies: static sampling, where the scores are determined before training, and dynamic sampling, where the scores can depend on the model weights. Static algorithms are computationally inexpensive but less effective than their dynamic counterparts, which can cause end-to-end slowdown due to their need to explicitly compute losses. To address this problem, we propose a novel sampling distribution based on nonparametric kernel regression that learns an effective importance score as the neural network trains. However, nonparametric regression models are too computationally expensive to accelerate end-to-end training. Therefore, we develop an efficient sketch-based approximation to the Nadaraya-Watson estimator. Using recent techniques from high-dimensional statistics and randomized algorithms, we prove that our Nadaraya-Watson sketch approximates the estimator with exponential convergence guarantees. Our sampling algorithm outperforms the baseline in terms of wall-clock time and accuracy on four datasets.

📄 PDF Abstract BibTeX arXiv:2311.13583

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Learningregression

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Prices, Profits, Proxies, and Production

2018-10-10 · Victor H. Aguiar, Nail Kashaev, Roy Allen

This paper studies nonparametric identification and counterfactual bounds for heterogeneous firms that can be ranked in terms of productivity. Our approach works when quantities and prices are latent, rendering standard …

counterfactual

Fewer is More: A Deep Graph Metric Learning Perspective Using Fewer Proxies

2020-10-26 · NeurIPS 2020 12 · Yuehua Zhu, Muli Yang, Cheng Deng, Wei Liu

Deep metric learning plays a key role in various machine learning tasks. Most of the previous works have been confined to sampling from a mini-batch, which cannot precisely characterize the global geometry of the embeddi…

General ClassificationGraph ClassificationMetric Learning

Proxy Controls and Panel Data

2018-09-30 · Ben Deaner

We provide new results for nonparametric identification, estimation, and inference of causal effects using `proxy controls': observables that are noisy but informative proxies for unobserved confounding factors. Our anal…

Adaptive and non-adaptive minimax rates for weighted Laplacian-eigenmap based nonparametric regression

2023-10-31 · Zhaoyang Shi, Krishnakumar Balasubramanian, Wolfgang Polonik

We show both adaptive and non-adaptive minimax rates of convergence for a family of weighted Laplacian-Eigenmap based nonparametric regression methods, when the true regression function belongs to a Sobolev space and the…

regression

Controlling for Latent Confounding with Triple Proxies

2022-04-28 · Ben Deaner

We present new results for nonparametric identification of causal effects using noisy proxies for unobserved confounders. Our approach builds on the results of \citet{Hu2008} who tackle the problem of general measurement…