paper-with-me

홈 › Papers

Structured Policy Representation: Imposing Stability in arbitrarily conditioned dynamic systems

2020-12-11 · Julen Urain, Davide Tateo, Tianyu Ren, Jan Peters

We present a new family of deep neural network-based dynamic systems. The presented dynamics are globally stable and can be conditioned with an arbitrary context state. We show how these dynamics can be used as structured robot policies. Global stability is one of the most important and straightforward inductive biases as it allows us to impose reasonable behaviors outside the region of the demonstrations.

📄 PDF Abstract BibTeX arXiv:2012.06224

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Latent Convergence Modulation in Large Language Models: A Novel Approach to Iterative Contextual Realignment

2025-02-10 · Patricia Porretta, Sylvester Pakenham, Huxley Ainsworth, Gregory Chatten 외

Token prediction stability remains a challenge in autoregressive generative models, where minor variations in early inference steps often lead to significant semantic drift over extended sequences. A structured modulatio…

Computational EfficiencySentenceText Generation

Imposing Robust Structured Control Constraint on Reinforcement Learning of Linear Quadratic Regulator

2020-11-12 · Sayak Mukherjee, Thanh Long Vu

This paper discusses learning a structured feedback control to obtain sufficient robustness to exogenous inputs for linear dynamic systems with unknown state matrix. The structural constraint on the controller is necessa…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Proactive Constrained Policy Optimization with Preemptive Penalty

2025-08-03 · Ning Yang, Pengyu Wang, Guoqing Liu, Haifeng Zhang 외 arxiv

Safe Reinforcement Learning (RL) often faces significant issues such as constraint violations and instability, necessitating the use of constrained policy optimization, which seeks optimal policies while ensuring adheren…

Reinforcement Learning

Demystifying Action Space Design for Robotic Manipulation Policies

2026-02-26 · Yuchun Feng, Jinliang Zheng, Zhihao Wang, Dongxiu Liu 외 arxiv

The specification of the action space plays a pivotal role in imitation-based robotic manipulation policy learning, fundamentally shaping the optimization landscape of policy learning. While recent advances have focused …

Stability-Constrained Markov Decision Processes Using MPC

2021-02-02 · Mario Zanon, Sébastien Gros, Michele Palladino

In this paper, we consider solving discounted Markov Decision Processes (MDPs) under the constraint that the resulting policy is stabilizing. In practice MDPs are solved based on some form of policy approximation. We wil…

Model Predictive Control