paper-with-me

홈 › Papers

An Upper Bound of the Bias of Nadaraya-Watson Kernel Regression under Lipschitz Assumptions

2020-01-29 · Samuele Tosatto, Riad Akrour, Jan Peters

The Nadaraya-Watson kernel estimator is among the most popular nonparameteric regression technique thanks to its simplicity. Its asymptotic bias has been studied by Rosenblatt in 1969 and has been reported in a number of related literature. However, Rosenblatt's analysis is only valid for infinitesimal bandwidth. In contrast, we propose in this paper an upper bound of the bias which holds for finite bandwidths. Moreover, contrarily to the classic analysis we allow for discontinuous first order derivative of the regression function, we extend our bounds for multidimensional domains and we include the knowledge of the bound of the regression function when it exists and if it is known, to obtain a tighter bound. We believe that this work has potential applications in those fields where some hard guarantees on the error are needed

📄 PDF Abstract BibTeX arXiv:2001.10972

Code (0)

등록된 구현이 없습니다.

Tasks

regressionvalid

Similar Papers 제목 키워드 기반

Heterogeneous Treatment Effect with Trained Kernels of the Nadaraya-Watson Regression

2022-07-19 · Andrei V. Konstantinov, Stanislav R. Kirpichenko, Lev V. Utkin

A new method for estimating the conditional average treatment effect is proposed in the paper. It is called TNW-CATE (the Trainable Nadaraya-Watson regression for CATE) and based on the assumption that the number of cont…

regressionTransfer Learning

Understanding the Mixture-of-Experts with Nadaraya-Watson Kernel

2025-09-30 · Chuanyang Zheng, Jiankai Sun, Yihang Gao, Enze Xie 외 arxiv

Mixture-of-Experts (MoE) has become a cornerstone in recent state-of-the-art large language models (LLMs). Traditionally, MoE relies on $\mathrm{Softmax}$ as the router score function to aggregate expert output, a design…

Nadaraya-Watson kernel smoothing as a random energy model

2024-08-07 · Jacob A. Zavatone-Veth, Cengiz Pehlevan

Precise asymptotics have revealed many surprises in high-dimensional regression. These advances, however, have not extended to perhaps the simplest estimator: direct Nadaraya-Watson (NW) kernel smoothing. Here, we descri…

Cubit: Token Mixer with Kernel Ridge Regression

2026-05-07 · Chuanyang Zheng, Jiankai Sun, Yihang Gao, Yuehao Wang 외 arxiv

Since its introduction in 2017, the Transformer has become one of the most widely adopted architectures in modern deep learning. Despite extensive efforts to improve positional encoding, attention mechanisms, and feed-fo…

Generative Local Metric Learning for Kernel Regression

2017-12-01 · NeurIPS 2017 12 · Yung-Kyun Noh, Masashi Sugiyama, Kee-Eung Kim, Frank Park 외

This paper shows how metric learning can be used with Nadaraya-Watson (NW) kernel regression. Compared with standard approaches, such as bandwidth selection, we show how metric learning can significantly reduce the mean…

Metric Learningregression