paper-with-me

홈 › Papers

Learning walk and trot from the same objective using different types of exploration

2019-04-28 · Zinan Liu, Kai Ploeger, Svenja Stark, Elmar Rueckert, Jan Peters

In quadruped gait learning, policy search methods that scale high dimensional continuous action spaces are commonly used. In most approaches, it is necessary to introduce prior knowledge on the gaits to limit the highly non-convex search space of the policies. In this work, we propose a new approach to encode the symmetry properties of the desired gaits, on the initial covariance of the Gaussian search distribution, allowing for strategic exploration. Using episode-based likelihood ratio policy gradient and relative entropy policy search, we learned the gaits walk and trot on a simulated quadruped. Comparing these gaits to random gaits learned by initialized diagonal covariance matrix, we show that the performance can be significantly enhanced.

📄 PDF Abstract BibTeX arXiv:1904.12336

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Fast and Efficient Locomotion via Learned Gait Transitions

2021-04-09 · Yuxiang Yang, Tingnan Zhang, Erwin Coumans, Jie Tan 외

We focus on the problem of developing energy efficient controllers for quadrupedal robots. Animals can actively switch gaits at different speeds to lower their energy consumption. In this paper, we devise a hierarchical …

Retrotransposon mobilization in cancer genomes

2015-01-18

The Cancer Genome Atlas project was initiated by the National Cancer Institute in order to characterize the genomes of hundreds of tumors of various cancer types. While much effort has been put into detecting somatic gen…

DeepTransition: Viability Leads to the Emergence of Gait Transitions in Learning Anticipatory Quadrupedal Locomotion Skills

2023-06-12 · Milad Shafiee, Guillaume Bellegarda, Auke Ijspeert

Quadruped animals seamlessly transition between gaits as they change locomotion speeds. While the most widely accepted explanation for gait transitions is energy efficiency, there is no clear consensus on the determining…

Deep Reinforcement Learning

Realizing Learned Quadruped Locomotion Behaviors through Kinematic Motion Primitives

2018-10-09 · Abhik Singla, Shounak Bhattacharya, Dhaivat Dholakiya, Shalabh Bhatnagar 외

Humans and animals are believed to use a very minimal set of trajectories to perform a wide variety of tasks including walking. Our main objective in this paper is two fold 1) Obtain an effective tool to realize these ba…

Deep Reinforcement LearningReinforcement Learning

OptRot: Mitigating Weight Outliers via Data-Free Rotations for Post-Training Quantization

2025-12-30 · Advait Gadhikar, Riccardo Grazzi, James Hensman arxiv

The presence of outliers in Large Language Models (LLMs) weights and activations makes them difficult to quantize. Recent work has leveraged rotations to mitigate these outliers. In this work, we propose methods that lea…