Learning Active Task-Oriented Exploration Policies for Bridging the Sim-to-Real Gap

Liang, Jacky; Saxena, Saumya; Kroemer, Oliver

Computer Science > Robotics

arXiv:2006.01952 (cs)

[Submitted on 2 Jun 2020 (v1), last revised 5 Nov 2020 (this version, v2)]

Title:Learning Active Task-Oriented Exploration Policies for Bridging the Sim-to-Real Gap

Authors:Jacky Liang, Saumya Saxena, Oliver Kroemer

View PDF

Abstract:Training robotic policies in simulation suffers from the sim-to-real gap, as simulated dynamics can be different from real-world dynamics. Past works tackled this problem through domain randomization and online system-identification. The former is sensitive to the manually-specified training distribution of dynamics parameters and can result in behaviors that are overly conservative. The latter requires learning policies that concurrently perform the task and generate useful trajectories for system identification. In this work, we propose and analyze a framework for learning exploration policies that explicitly perform task-oriented exploration actions to identify task-relevant system parameters. These parameters are then used by model-based trajectory optimization algorithms to perform the task in the real world. We instantiate the framework in simulation with the Linear Quadratic Regulator as well as in the real world with pouring and object dragging tasks. Experiments show that task-oriented exploration helps model-based policies adapt to systems with initially unknown parameters, and it leads to better task performance than task-agnostic exploration.

Comments:	Published at Robotics: Science and Systems 2020
Subjects:	Robotics (cs.RO)
Cite as:	arXiv:2006.01952 [cs.RO]
	(or arXiv:2006.01952v2 [cs.RO] for this version)
	https://doi.org/10.48550/arXiv.2006.01952

Submission history

From: Jacky Liang [view email]
[v1] Tue, 2 Jun 2020 21:32:58 UTC (5,557 KB)
[v2] Thu, 5 Nov 2020 22:05:53 UTC (8,573 KB)

Computer Science > Robotics

Title:Learning Active Task-Oriented Exploration Policies for Bridging the Sim-to-Real Gap

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Robotics

Title:Learning Active Task-Oriented Exploration Policies for Bridging the Sim-to-Real Gap

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators