NFDI4DS | UHH-SEMS - Publication Details

no need for interactions robust model based imitation learning using neural ode

FOS: Computer and information sciences Computer Science - Robotics Computer Science - Machine Learning FOS: Electrical engineering, electronic engineering, information engineering 0202 electrical engineering, electronic engineering, information engineering Systems and Control (eess.SY) 02 engineering and technology Electrical Engineering and Systems Science - Systems and Control Robotics (cs.RO) Machine Learning (cs.LG)

DOI: 10.48550/arxiv.2104.01390 Publication Date: 2021-05-30

Abstract Supplemental Material References Cited by

AUTHORS (5)

Baopu Li

Max Q.-H. Meng

HaoChih Lin

Xin Zhou

Jiankun Wang

ABSTRACT

Interactions with either environments or expert policies during training are needed for most of the current imitation learning (IL) algorithms. For IL problems with no interactions, a typical approach is Behavior Cloning (BC). However, BC-like methods tend to be affected by distribution shift. To mitigate this problem, we come up with a Robust Model-Based Imitation Learning (RMBIL) framework that casts imitation learning as an end-to-end differentiable nonlinear closed-loop tracking problem. RMBIL applies Neural ODE to learn a precise multi-step dynamics and a robust tracking controller via Nonlinear Dynamics Inversion (NDI) algorithm. Then, the learned NDI controller will be combined with a trajectory generator, a conditional VAE, to imitate an expert's behavior. Theoretical derivation shows that the controller network can approximate an NDI when minimizing the training loss of Neural ODE. Experiments on Mujoco tasks also demonstrate that RMBIL is competitive to the state-of-the-art generative adversarial method (GAIL) and achieves at least 30% performance gain over BC in uneven surfaces.

SUPPLEMENTAL MATERIAL

Coming soon ....

REFERENCES ()

CITATIONS ()

EXTERNAL LINKS

OPENAIRE - Products

PlumX Metrics

no need for interactions robust model based imitation learning using neural ode

RECOMMENDATIONS

FAIR ASSESSMENT

Coming soon ....

JUPYTER LAB

Coming soon ....