paper-with-me

Papers

Wasserstein Distance Maximizing Intrinsic Control

2021-10-28 · Ishan Durugkar, Steven Hansen, Stephen Spencer, Volodymyr Mnih

This paper deals with the problem of learning a skill-conditioned policy that acts meaningfully in the absence of a reward signal. Mutual information based objectives have shown some success in learning skills that reach a diverse set of states in this setting. These objectives include a KL-divergence term, which is maximized by visiting distinct states even if those states are not far apart in the MDP. This paper presents an approach that rewards the agent for learning skills that maximize the Wasserstein distance of their state visitation from the start state of the skill. It shows that such an objective leads to a policy that covers more distance in the MDP than diversity based objectives, and validates the results on a variety of Atari environments.

📄 PDF Abstract BibTeX arXiv:2110.15331

Code (0)

등록된 구현이 없습니다.

Tasks

Diversity

Similar Papers 제목 키워드 기반

Shape Analysis With Hyperbolic Wasserstein Distance

2016-06-01 · CVPR 2016 6 · Jie Shi, Wen Zhang, Yalin Wang

Shape space is an active research field in computer vision study. The shape distance defined in a shape space may provide a simple and refined index to represent a unique shape. Wasserstein distance defines a Riemannian …

ClassificationGeneral Classification

Geometry-Aware Optimal Transport: Fast Intrinsic Dimension and Wasserstein Distance Estimation

2026-02-04 · Ferdinand Genans, Olivier Wintenberger arxiv

Solving large scale Optimal Transport (OT) in machine learning typically relies on sampling measures to obtain a tractable discrete problem. While the discrete solver's accuracy is controllable, the rate of convergence o…

Computational Efficiency

Distance-Matrix Wasserstein Statistics for Scalable Gromov--Wasserstein Learning

2026-05-14 · Ao Xu, Tieru Wu arxiv

Gromov--Wasserstein (GW) distances compare graphs, shapes, and point clouds through internal distances, without requiring a common coordinate system. This invariance is powerful, but discrete GW is a nonconvex quadratic …

Graph ClassificationTwo-sample testingPoint Clouds

Feature Selection via Maximizing Distances between Class Conditional Distributions

2024-01-15 · Chunxu Cao, Qiang Zhang

For many data-intensive tasks, feature selection is an important preprocessing step. However, most existing methods do not directly and intuitively explore the intrinsic discriminative information of features. We propose…

feature selection

Wasserstein Nonnegative Tensor Factorization with Manifold Regularization

2024-01-03 · Jianyu Wang, Linruize Tang

Nonnegative tensor factorization (NTF) has become an important tool for feature extraction and part-based representation with preserved intrinsic structure information from nonnegative high-order data. However, the origi…