paper-with-me

Papers

Dual Advantage Fields

2026-06-02 · Alexey Zemtsov, Maxim Bobrin, Alexander Nikulin, Dmitry V. Dylov, Fakhri Karray, Vladislav Kurenkov, Martin Takáč, Arip Asadulaev arxiv

Offline goal-conditioned reinforcement learning requires both long-horizon reachability estimates and local action comparisons. Dual goal representations provide value fields that capture global goal reachability, but they do not directly specify which action should be preferred at a given state. We propose Dual Advantage Fields, a policy-extraction method that turns a bilinear dual value model into a local advantage signal. Under bilinear dual parameterization, the goal embedding is the gradient of the value field with respect to the state representation. DAF learns an action-effect model that predicts the discounted feature displacement induced by an action and scores actions by the alignment between this displacement and the goal direction. In the realizable case, this score equals the goal-conditioned Bellman advantage, yielding a standard local policy-improvement guarantee. On OGBench locomotion, manipulation, and puzzle tasks, DAF improves aggregate RLiable metrics and performs strongly in settings where locally correct actions differ from direct movement toward the final goal.

📄 PDF Abstract BibTeX arXiv:2606.04188

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Depth Fields: Extending Light Field Techniques to Time-of-Flight Imaging

2015-09-02 · Suren Jayasuriya, Adithya Pediredla, Sriram Sivaramakrishnan, Alyosha Molnar 외

A variety of techniques such as light field, structured illumination, and time-of-flight (TOF) are commonly used for depth acquisition in consumer imaging, robotics and many other applications. Unfortunately, each techni…

Human-Machine Cooperative Multimodal Learning Method for Cross-subject Olfactory Preference Recognition

2023-11-24 · Xiuxin Xia, Yuchen Guo, Yanwei Wang, Yuchao Yang 외

Odor sensory evaluation has a broad application in food, clothing, cosmetics, and other fields. Traditional artificial sensory evaluation has poor repeatability, and the machine olfaction represented by the electronic no…

EEGElectroencephalogram (EEG)

Neural Diffeomorphic-Neural Operator for Residual Stress-Induced Deformation Prediction

2025-09-09 · Changqing Liu, Kaining Dai, Zhiwei Zhao, Tianyi Wu 외 arxiv

Accurate prediction of machining deformation in structural components is essential for ensuring dimensional precision and reliability. Such deformation often originates from residual stress fields, whose distribution and…

Canny-VO: Visual Odometry with RGB-D Cameras based on Geometric 3D-2D Edge Alignment

2020-12-15 · Yi Zhou, Hongdong Li, Laurent Kneip

The present paper reviews the classical problem of free-form curve registration and applies it to an efficient RGBD visual odometry system called Canny-VO, as it efficiently tracks all Canny edge features extracted from …

Visual Odometry

Applications of Recurrent Neural Network for Biometric Authentication & Anomaly Detection

2021-09-13 · Joseph M. Ackerson, Dave Rushit, Seliya Jim

Recurrent Neural Networks are powerful machine learning frameworks that allow for data to be saved and referenced in a temporal sequence. This opens many new possibilities in fields such as handwriting analysis and speec…

Anomaly Detectionspeech-recognitionSpeech Recognition