paper-with-me

Papers

Model-Based Reinforcement Learning for Physical Systems Without Velocity and Acceleration Measurements

2020-02-25 · Alberto Dalla Libera, Diego Romeres, Devesh K. Jha, Bill Yerazunis, Daniel Nikovski

In this paper, we propose a derivative-free model learning framework for Reinforcement Learning (RL) algorithms based on Gaussian Process Regression (GPR). In many mechanical systems, only positions can be measured by the sensing instruments. Then, instead of representing the system state as suggested by the physics with a collection of positions, velocities, and accelerations, we define the state as the set of past position measurements. However, the equation of motions derived by physical first principles cannot be directly applied in this framework, being functions of velocities and accelerations. For this reason, we introduce a novel derivative-free physically-inspired kernel, which can be easily combined with nonparametric derivative-free Gaussian Process models. Tests performed on two real platforms show that the considered state definition combined with the proposed model improves estimation performance and data-efficiency w.r.t. traditional models based on GPR. Finally, we validate the proposed framework by solving two RL control problems for two real robotic systems.

📄 PDF Abstract BibTeX arXiv:2002.10621

Code (0)

등록된 구현이 없습니다.

Tasks

GPRModel-based Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Methods 이 논문이 사용한 방법론

Gaussian Process Gaussian Processes are non-parametric models for approximating functions. They rely upon a measure of similarity between points (the kernel function) to predict the value for…

Similar Papers 제목 키워드 기반

Learning Velocity and Acceleration: Self-Supervised Motion Consistency for Pedestrian Trajectory Prediction

2025-03-31 · Yizhou Huang, Yihua Cheng, Kezhi Wang

Understanding human motion is crucial for accurate pedestrian trajectory prediction. Conventional methods typically rely on supervised learning, where ground-truth labels are directly optimized against predicted trajecto…

Pedestrian Trajectory PredictionPositionSelf-Supervised LearningTrajectory Prediction

OAT-FM: Optimal Acceleration Transport for Improved Flow Matching

2025-09-29 · Angxiao Yue, Anqi Dong, Hongteng Xu arxiv

As a powerful technique in generative modeling, Flow Matching (FM) aims to learn velocity fields from noise to data, which is often explained and implemented as solving Optimal Transport (OT) problems. In this study, we …

VRA: Grounding Discrete-Time Joint Acceleration in Voltage-Constrained Actuation

2026-05-11 · Lingwei Zhang, Jiaming Wang, Tianlin Zhang, Zhitao Song 외 arxiv

Discrete-time joint acceleration constraints are widely used to enforce position and velocity limits. However, under voltage-constrained electric actuators, kinematically admissible accelerations may be physically unreal…

Angular velocity and linear acceleration measurement bias estimators for the rigid body system with global exponential convergence

2024-01-20 · Soham Shanbhag, Dong Eui Chang

Rigid body systems usually consider measurements of the pose of the body using onboard cameras/LiDAR systems, that of linear acceleration using an accelerometer and of angular velocity using an IMU. However, the measurem…

Regularizing Dynamic Radiance Fields with Kinematic Fields

2024-07-19 · Woobin Im, Geonho Cha, Sebin Lee, Jumin Lee 외

This paper presents a novel approach for reconstructing dynamic radiance fields from monocular videos. We integrate kinematics with dynamic radiance fields, bridging the gap between the sparse nature of monocular videos …