Open library

This portal has been archived. Explore the next generation of this technology.

Learning walk and trot from the same objective using different types of exploration

lib:e31464a56a2289fa (v1.0.0)

Authors: Zinan Liu,Kai Ploeger,Svenja Stark,Elmar Rueckert,Jan Peters
ArXiv: 1904.12336
Document: PDF DOI

Abstract URL: http://arxiv.org/abs/1904.12336v1

In quadruped gait learning, policy search methods that scale high dimensional continuous action spaces are commonly used. In most approaches, it is necessary to introduce prior knowledge on the gaits to limit the highly non-convex search space of the policies. In this work, we propose a new approach to encode the symmetry properties of the desired gaits, on the initial covariance of the Gaussian search distribution, allowing for strategic exploration. Using episode-based likelihood ratio policy gradient and relative entropy policy search, we learned the gaits walk and trot on a simulated quadruped. Comparing these gaits to random gaits learned by initialized diagonal covariance matrix, we show that the performance can be significantly enhanced.

Relevant initiatives

Related knowledge about this paper

Search on this portal

Reproduced results (crowd-benchmarking and competitions)

Artifact and reproducibility checklists

Common formats for research projects and shared artifacts

Collective Knowledge (organizing research projects based on FAIR principles)

Reproducibility initiatives

Comments

Please log in to add your comments!

If you notice any inapropriate content that should not be here, please report us as soon as possible and we will try to remove it within 48 hours!

Learning walk and trot from the same objective using different types of exploration

Relevant initiatives Hide

Comments Hide

Relevant initiatives

Comments