Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Contextual Latent-Movements Off-Policy Optimization for Robotic Manipulation Skills

Oct 26, 2020

Samuele Tosatto, Georgia Chalvatzaki, Jan Peters

Figure 1 for Contextual Latent-Movements Off-Policy Optimization for Robotic Manipulation Skills

Figure 2 for Contextual Latent-Movements Off-Policy Optimization for Robotic Manipulation Skills

Figure 3 for Contextual Latent-Movements Off-Policy Optimization for Robotic Manipulation Skills

Figure 4 for Contextual Latent-Movements Off-Policy Optimization for Robotic Manipulation Skills

Share this with someone who'll enjoy it:

Abstract:Parameterized movement primitives have been extensively used for imitation learning of robotic tasks. However, the high-dimensionality of the parameter space hinders the improvement of such primitives in the reinforcement learning (RL) setting, especially for learning with physical robots. In this paper we propose a novel view on handling the demonstrated trajectories for acquiring low-dimensional, non-linear latent dynamics, using mixtures of probabilistic principal component analyzers (MPPCA) on the movements' parameter space. Moreover, we introduce a new contextual off-policy RL algorithm, named LAtent-Movements Policy Optimization (LAMPO). LAMPO can provide gradient estimates from previous experience using self-normalized importance sampling, hence, making full use of samples collected in previous learning iterations. These advantages combined provide a complete framework for sample-efficient off-policy optimization of movement primitives for robot learning of high-dimensional manipulation skills. Our experimental results conducted both in simulation and on a real robot show that LAMPO provides sample-efficient policies against common approaches in literature.

View paper on

Share this with someone who'll enjoy it:

Title:Contextual Latent-Movements Off-Policy Optimization for Robotic Manipulation Skills

Paper and Code