'Good Robot!': Efficient Reinforcement Learning for Multi-Step Visual Tasks with Sim to Real Transfer

Andrew Hundt; Benjamin Killeen; Nicholas Greene; Hongtao Wu; Heeyeon Kwon; Chris Paxton; Gregory D. Hager

doi:10.1109/LRA.2020.3015448

'Good Robot!': Efficient Reinforcement Learning for Multi-Step Visual Tasks with Sim to Real Transfer

Andrew Hundt, Benjamin Killeen, Nicholas Greene, Hongtao Wu, Heeyeon Kwon, Chris Paxton, Gregory D. Hager

Whiting School of Engineering

Research output: Contribution to journal › Article › peer-review

2 Scopus citations

Abstract

Current Reinforcement Learning (RL) algorithms struggle with long-horizon tasks where time can be wasted exploring dead ends and task progress may be easily reversed. We develop the SPOT framework, which explores within action safety zones, learns about unsafe regions without exploring them, and prioritizes experiences that reverse earlier progress to learn with remarkable efficiency. The SPOT framework successfully completes simulated trials of a variety of tasks, improving a baseline trial success rate from 13% to 100% when stacking 4 cubes, from 13% to 99% when creating rows of 4 cubes, and from 84% to 95% when clearing toys arranged in adversarial patterns. Efficiency with respect to actions per trial typically improves by 30% or more, while training takes just 1-20 k actions, depending on the task. Furthermore, we demonstrate direct sim to real transfer. We are able to create real stacks in 100% of trials with 61% efficiency and real rows in 100% of trials with 59% efficiency by directly loading the simulation-trained model on the real robot with no additional real-world fine-tuning. To our knowledge, this is the first instance of reinforcement learning with successful sim to real transfer applied to long term multi-step tasks such as block-stacking and row-making with consideration of progress reversal. Code is available at https://github.com/jhu-lcsr/good_robot.

Original language	English (US)
Article number	9165109
Pages (from-to)	6724-6731
Number of pages	8
Journal	IEEE Robotics and Automation Letters
Volume	5
Issue number	4
DOIs	https://doi.org/10.1109/LRA.2020.3015448
State	Published - Oct 2020

Keywords

Computer vision for other robotic applications
deep learning in grasping and manipulation
reinforcement learning

ASJC Scopus subject areas

Control and Systems Engineering
Biomedical Engineering
Human-Computer Interaction
Mechanical Engineering
Computer Vision and Pattern Recognition
Computer Science Applications
Control and Optimization
Artificial Intelligence

Access to Document

10.1109/LRA.2020.3015448

Cite this

@article{6705ca21a0fc48c1a7cc05e578127757,

title = "'Good Robot!': Efficient Reinforcement Learning for Multi-Step Visual Tasks with Sim to Real Transfer",

abstract = "Current Reinforcement Learning (RL) algorithms struggle with long-horizon tasks where time can be wasted exploring dead ends and task progress may be easily reversed. We develop the SPOT framework, which explores within action safety zones, learns about unsafe regions without exploring them, and prioritizes experiences that reverse earlier progress to learn with remarkable efficiency. The SPOT framework successfully completes simulated trials of a variety of tasks, improving a baseline trial success rate from 13% to 100% when stacking 4 cubes, from 13% to 99% when creating rows of 4 cubes, and from 84% to 95% when clearing toys arranged in adversarial patterns. Efficiency with respect to actions per trial typically improves by 30% or more, while training takes just 1-20 k actions, depending on the task. Furthermore, we demonstrate direct sim to real transfer. We are able to create real stacks in 100% of trials with 61% efficiency and real rows in 100% of trials with 59% efficiency by directly loading the simulation-trained model on the real robot with no additional real-world fine-tuning. To our knowledge, this is the first instance of reinforcement learning with successful sim to real transfer applied to long term multi-step tasks such as block-stacking and row-making with consideration of progress reversal. Code is available at https://github.com/jhu-lcsr/good_robot.",

keywords = "Computer vision for other robotic applications, deep learning in grasping and manipulation, reinforcement learning",

author = "Andrew Hundt and Benjamin Killeen and Nicholas Greene and Hongtao Wu and Heeyeon Kwon and Chris Paxton and Hager, {Gregory D.}",

note = "Publisher Copyright: {\textcopyright} 2016 IEEE.",

year = "2020",

month = oct,

doi = "10.1109/LRA.2020.3015448",

language = "English (US)",

volume = "5",

pages = "6724--6731",

journal = "IEEE Robotics and Automation Letters",

issn = "2377-3766",

publisher = "Institute of Electrical and Electronics Engineers Inc.",

number = "4",

}

TY - JOUR

T1 - 'Good Robot!'

T2 - Efficient Reinforcement Learning for Multi-Step Visual Tasks with Sim to Real Transfer

AU - Hundt, Andrew

AU - Killeen, Benjamin

AU - Greene, Nicholas

AU - Wu, Hongtao

AU - Kwon, Heeyeon

AU - Paxton, Chris

AU - Hager, Gregory D.

PY - 2020/10

Y1 - 2020/10

N2 - Current Reinforcement Learning (RL) algorithms struggle with long-horizon tasks where time can be wasted exploring dead ends and task progress may be easily reversed. We develop the SPOT framework, which explores within action safety zones, learns about unsafe regions without exploring them, and prioritizes experiences that reverse earlier progress to learn with remarkable efficiency. The SPOT framework successfully completes simulated trials of a variety of tasks, improving a baseline trial success rate from 13% to 100% when stacking 4 cubes, from 13% to 99% when creating rows of 4 cubes, and from 84% to 95% when clearing toys arranged in adversarial patterns. Efficiency with respect to actions per trial typically improves by 30% or more, while training takes just 1-20 k actions, depending on the task. Furthermore, we demonstrate direct sim to real transfer. We are able to create real stacks in 100% of trials with 61% efficiency and real rows in 100% of trials with 59% efficiency by directly loading the simulation-trained model on the real robot with no additional real-world fine-tuning. To our knowledge, this is the first instance of reinforcement learning with successful sim to real transfer applied to long term multi-step tasks such as block-stacking and row-making with consideration of progress reversal. Code is available at https://github.com/jhu-lcsr/good_robot.

AB - Current Reinforcement Learning (RL) algorithms struggle with long-horizon tasks where time can be wasted exploring dead ends and task progress may be easily reversed. We develop the SPOT framework, which explores within action safety zones, learns about unsafe regions without exploring them, and prioritizes experiences that reverse earlier progress to learn with remarkable efficiency. The SPOT framework successfully completes simulated trials of a variety of tasks, improving a baseline trial success rate from 13% to 100% when stacking 4 cubes, from 13% to 99% when creating rows of 4 cubes, and from 84% to 95% when clearing toys arranged in adversarial patterns. Efficiency with respect to actions per trial typically improves by 30% or more, while training takes just 1-20 k actions, depending on the task. Furthermore, we demonstrate direct sim to real transfer. We are able to create real stacks in 100% of trials with 61% efficiency and real rows in 100% of trials with 59% efficiency by directly loading the simulation-trained model on the real robot with no additional real-world fine-tuning. To our knowledge, this is the first instance of reinforcement learning with successful sim to real transfer applied to long term multi-step tasks such as block-stacking and row-making with consideration of progress reversal. Code is available at https://github.com/jhu-lcsr/good_robot.

KW - Computer vision for other robotic applications

KW - deep learning in grasping and manipulation

KW - reinforcement learning

UR - http://www.scopus.com/inward/record.url?scp=85089452039&partnerID=8YFLogxK

UR - http://www.scopus.com/inward/citedby.url?scp=85089452039&partnerID=8YFLogxK

U2 - 10.1109/LRA.2020.3015448

DO - 10.1109/LRA.2020.3015448

M3 - Article

AN - SCOPUS:85089452039

SN - 2377-3766

VL - 5

SP - 6724

EP - 6731

JO - IEEE Robotics and Automation Letters

JF - IEEE Robotics and Automation Letters

IS - 4

M1 - 9165109

ER -

'Good Robot!': Efficient Reinforcement Learning for Multi-Step Visual Tasks with Sim to Real Transfer

Abstract

Keywords

ASJC Scopus subject areas

Access to Document

Other files and links

Fingerprint

Cite this