paper-with-me

홈 › Papers

Canonical Variates in Wasserstein Metric Space

2024-05-24 · Jia Li, Lin Lin

In this paper, we address the classification of instances each characterized not by a singular point, but by a distribution on a vector space. We employ the Wasserstein metric to measure distances between distributions, which are then used by distance-based classification algorithms such as k-nearest neighbors, k-means, and pseudo-mixture modeling. Central to our investigation is dimension reduction within the Wasserstein metric space to enhance classification accuracy. We introduce a novel approach grounded in the principle of maximizing Fisher's ratio, defined as the quotient of between-class variation to within-class variation. The directions in which this ratio is maximized are termed discriminant coordinates or canonical variates axes. In practice, we define both between-class and within-class variations as the average squared distances between pairs of instances, with the pairs either belonging to the same class or to different classes. This ratio optimization is achieved through an iterative algorithm, which alternates between optimal transport and maximization steps within the vector space. We conduct empirical studies to assess the algorithm's convergence and, through experimental validation, demonstrate that our dimension reduction technique substantially enhances classification performance. Moreover, our method outperforms well-established algorithms that operate on vector representations derived from distributional data. It also exhibits robustness against variations in the distributional representations of data clouds.

📄 PDF Abstract BibTeX arXiv:2405.15768

Code (0)

등록된 구현이 없습니다.

Tasks

Dimensionality Reduction

Similar Papers 제목 키워드 기반

Subspace Perspective on Canonical Correlation Analysis: Dimension Reduction and Minimax Rates

2016-05-12 · Zhuang Ma, Xiao-Dong Li

Canonical correlation analysis (CCA) is a fundamental statistical tool for exploring the correlation structure between two sets of random variables. In this paper, motivated by recent success of applying CCA to learn low…

Dimensionality ReductionPrediction

Bounds in Wasserstein Distance for Locally Stationary Functional Time Series

2025-04-08 · Jan Nino G. Tinio, Mokhtar Z. Alaya, Salim Bouzebda

Functional time series (FTS) extend traditional methodologies to accommodate data observed as functions/curves. A significant challenge in FTS consists of accurately capturing the time-dependence structure, especially wi…

Time Series

Implicit Bias of the JKO Scheme

2025-11-18 · Peter Halmos, Boris Hanin arxiv

Wasserstein gradient flow provides a general framework for minimizing an energy functional $J$ over the space of probability measures on a Riemannian manifold $(M,g)$. Its canonical time-discretization, the Jordan-Kinder…

Revisiting Counterfactual Regression through the Lens of Gromov-Wasserstein Information Bottleneck

2024-05-24 · Hao Yang, Zexu Sun, Hongteng Xu, Xu Chen

As a promising individualized treatment effect (ITE) estimation method, counterfactual regression (CFR) maps individuals' covariates to a latent space and predicts their counterfactual outcomes. However, the selection bi…

counterfactualregressionSelection bias

Contextual Optimization under Covariate Shift: A Robust Approach by Intersecting Wasserstein Balls

2024-06-04 · Tianyu Wang, Ningyuan Chen, Chun Wang

In contextual optimization, a decision-maker leverages contextual information, often referred to as covariates, to better resolve uncertainty and make informed decisions. In this paper, we examine the challenges of conte…

Computational EfficiencyPortfolio Optimization