paper-with-me

Papers

On Projection Robust Optimal Transport: Sample Complexity and Model Misspecification

2020-06-22 · Tianyi Lin, Zeyu Zheng, Elynn Y. Chen, Marco Cuturi, Michael. I. Jordan

Optimal transport (OT) distances are increasingly used as loss functions for statistical inference, notably in the learning of generative models or supervised learning. Yet, the behavior of minimum Wasserstein estimators is poorly understood, notably in high-dimensional regimes or under model misspecification. In this work we adopt the viewpoint of projection robust (PR) OT, which seeks to maximize the OT cost between two measures by choosing a $k$-dimensional subspace onto which they can be projected. Our first contribution is to establish several fundamental statistical properties of PR Wasserstein distances, complementing and improving previous literature that has been restricted to one-dimensional and well-specified cases. Next, we propose the integral PR Wasserstein (IPRW) distance as an alternative to the PRW distance, by averaging rather than optimizing on subspaces. Our complexity bounds can help explain why both PRW and IPRW distances outperform Wasserstein distances empirically in high-dimensional inference tasks. Finally, we consider parametric inference using the PRW distance. We provide an asymptotic guarantee of two types of minimum PRW estimators and formulate a central limit theorem for max-sliced Wasserstein estimator under model misspecification. To enable our analysis on PRW with projection dimension larger than one, we devise a novel combination of variational analysis and statistical theory.

📄 PDF Abstract BibTeX arXiv:2006.12301

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Statistical Optimal Transport posed as Learning Kernel Embedding

2020-02-08 · NeurIPS 2020 12 · J. Saketha Nath, Pratik Jawanpuria

The objective in statistical Optimal Transport (OT) is to consistently estimate the optimal transport plan/map solely using samples from the given source and target marginal distributions. This work takes the novel appro…

Theoretical guarantees for EM under misspecified Gaussian mixture models

2018-12-01 · NeurIPS 2018 12 · Raaz Dwivedi, Nhật Hồ, Koulik Khamaru, Martin J. Wainwright 외

Recent years have witnessed substantial progress in understanding the behavior of EM for mixture models that are correctly specified. Given that model misspecification is common in practice, it is important to unde…

parameter estimation

Lightspeed Geometric Dataset Distance via Sliced Optimal Transport

2025-01-31 · Khai Nguyen, Hai Nguyen, Tuan Pham, Nhat Ho

We introduce sliced optimal transport dataset distance (s-OTDD), a model-agnostic, embedding-agnostic approach for dataset comparison that requires no training, is robust to variations in the number of classes, and can h…

Data AugmentationTransfer Learning

Wasserstein k-means with sparse simplex projection

2020-11-25 · Takumi Fukunaga, Hiroyuki Kasai

This paper presents a proposal of a faster Wasserstein $k$-means algorithm for histogram data by reducing Wasserstein distance computations and exploiting sparse simplex projection. We shrink data samples, centroids, and…

Clustering

Sliced Multi-Marginal Optimal Transport

2021-02-14 · samuel cohen, Alexander Terenin, Yannik Pitcan, Brandon Amos 외

Multi-marginal optimal transport enables one to compare multiple probability measures, which increasingly finds application in multi-task learning problems. One practical limitation of multi-marginal transport is computa…

Density EstimationMulti-Task Learning