paper-with-me

Papers

QQWorld: Quantile-Quantile Matching for World Model Regularization

2026-07-30 · Zhoushun Yu, Xiaoyu Hu, Xiangyu Xu arxiv

Latent world models enable efficient planning by predicting future states in a compact representation space, but their performance depends critically on the quality of the learned latent distribution. LeWorldModel (LeWM) regularizes its latents toward an isotropic Gaussian using the Epps-Pulley (EP) objective. We show that the corrective gradients of EP rapidly vanish for isolated tail samples, leaving heavy-tailed deviations insufficiently controlled. To address this limitation, we propose QQWorld, which replaces EP with a quantile-quantile matching objective that directly aligns projected latent samples with rank-matched Gaussian quantiles, thereby maintaining effective corrective gradients in the tails. We further develop cross-batch QQ, which enlarges the effective ranking pool using detached samples from previous batches, and characterize its bias-variance trade-off. Across four control environments, QQWorld effectively improves the average planning success rate of LeWM, while consistently yielding better Gaussian alignment and thinner latent tails.

📄 PDF Abstract BibTeX arXiv:2607.28415

Code (2)

20bytes/Aerial-VLN-Arxiv-Daily ★ 8
BaiShuanghao/my_arXiv_daily ★ 205

Similar Papers 제목 키워드 기반

Distributions of Posterior Quantiles via Matching

2024-02-27 · Anton Kolotilin, Alexander Wolitzky

We offer a simple analysis of the problem of choosing a statistical experiment to optimize the induced distribution of posterior medians, or more generally $q$-quantiles for any $q \in (0,1)$. We show that all implementa…

Dataset Condensation with Latent Quantile Matching

2024-06-14 · Wei Wei, Tom De Schepper, Kevin Mets

Dataset condensation (DC) methods aim to learn a smaller synthesized dataset with informative data records to accelerate the training of machine learning models. Current distribution matching (DM) based DC methods learn …

Dataset CondensationGraph Learning

Regularization Strategies for Quantile Regression

2021-02-09 · Taman Narayan, Serena Wang, Kevin Canini, Maya Gupta

We investigate different methods for regularizing quantile regression when predicting either a subset of quantiles or the full inverse CDF. We show that minimizing an expected pinball loss over a continuous distribution …

Fairnessquantile regressionregression

Quantile Geometry Regularization for Distributional Reinforcement Learning

2026-05-05 · Zhaofan Zhang, Minghao Yang, Rufeng Chen, Sihong Xie 외 arxiv

Quantile-based distributional reinforcement learning methods learn return distributions through sampled quantile regression, but their bootstrapped target quantiles may induce distorted or degenerate distribution estimat…

Reinforcement LearningAtari Games

Bayesian Quantile Matching Estimation

2020-08-14 · Rajbir-Singh Nirwan, Nils Bertschinger

Due to increased awareness of data protection and corresponding laws many data, especially involving sensitive personal information, are not publicly accessible. Accordingly, many data collecting agencies only release ag…