paper-with-me

홈 › Papers

Beyond Fixed Morphologies: Learning Graph Policies with Trust Region Compensation in Variable Action Spaces

2025-08-16 · Thomas Gallien arxiv

Trust region-based optimization methods have become foundational reinforcement learning algorithms that offer stability and strong empirical performance in continuous control tasks. Growing interest in scalable and reusable control policies translate also in a demand for morphological generalization, the ability of control policies to cope with different kinematic structures. Graph-based policy architectures provide a natural and effective mechanism to encode such structural differences. However, while these architectures accommodate variable morphologies, the behavior of trust region methods under varying action space dimensionality remains poorly understood. To this end, we conduct a theoretical analysis of trust region-based policy optimization methods, focusing on both Trust Region Policy Optimization (TRPO) and its widely used first-order approximation, Proximal Policy Optimization (PPO). The goal is to demonstrate how varying action space dimensionality influence the optimization landscape, particularly under the constraints imposed by KL-divergence or policy clipping penalties. Complementing the theoretical insights, an empirical evaluation under morphological variation is carried out using the Gymnasium Swimmer environment. This benchmark offers a systematically controlled setting for varying the kinematic structure without altering the underlying task, making it particularly well-suited to study morphological generalization.

📄 PDF Abstract BibTeX arXiv:2508.14102

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement LearningContinuous Control

Similar Papers 제목 키워드 기반

UniBYD: A Unified Framework for Learning Robotic Manipulation Across Embodiments Beyond Imitation of Human Demonstrations

2025-12-12 · Tingyu Yuan, Biaoliang Guan, Wen Ye, Ziyan Tian 외 arxiv

In embodied intelligence, the embodiment gap between robotic and human hands brings significant challenges for learning from human demonstrations. Although some studies have attempted to bridge this gap using reinforceme…

Reinforcement Learning

Motion Beyond Morphology: Bootstrapping Cross-Category Motion Transfer from Abstract Motion Representations

2026-08-03 · Zhixue Fang, Zhimin Zhang, Bi'an Du, Zijie Meng 외 arxiv

Video motion transfer aims to animate a target object using dynamics from a reference video. Existing formulations largely rely on fixed structural correspondence, which becomes ill-defined when reference and target obje…

HeteroMorpheus: Universal Control Based on Morphological Heterogeneity Modeling

2024-08-02 · Yifan Hao, Yang Yang, Junru Song, Wei Peng 외

In the field of robotic control, designing individual controllers for each robot leads to high computational costs. Universal control policies, applicable across diverse robot morphologies, promise to mitigate this chall…

DiversityZero-shot Generalization

X-Morph: Human Motion Priors for Scalable Robot Learning Across Morphologies

2026-06-29 · Ritwik Sharma, Shivam Sood, Arhaan Jain, Shyam Charan Kesavamoorthi 외 arxiv

Recent progress in humanoid behavior models has been driven in large part by abundant human motion data, but comparable motion data is scarce for non-humanoid legged robots such as quadrupeds, hexapods, and quadruped man…

Universal Morphology Control via Contextual Modulation

2023-02-22 · Zheng Xiong, Jacob Beck, Shimon Whiteson

Learning a universal policy across different robot morphologies can significantly improve learning efficiency and generalization in continuous control. However, it poses a challenging multi-task reinforcement learning pr…

continuous-controlContinuous Control