OGMP: Oracle Guided Multi-mode Policies for Agile and Versatile Robot Control

Krishna, Lokesh; Sobanbabu, Nikhil; Nguyen, Quan

Computer Science > Robotics

arXiv:2403.04205 (cs)

[Submitted on 7 Mar 2024 (v1), last revised 18 Sep 2024 (this version, v3)]

Title:OGMP: Oracle Guided Multi-mode Policies for Agile and Versatile Robot Control

Authors:Lokesh Krishna, Nikhil Sobanbabu, Quan Nguyen

View PDF HTML (experimental)

Abstract:The efficacy of reinforcement learning for robot control relies on the tailored integration of task-specific priors and heuristics for effective exploration, which challenges their straightforward application to complex tasks and necessitates a unified approach. In this work, we define a general class for priors called oracles that generate state references when queried in a closed-loop manner during training. By bounding the permissible state around the oracle's ansatz, we propose a task-agnostic oracle-guided policy optimization. To enhance modularity, we introduce task-vital modes, showing that a policy mastering a compact set of modes and transitions can handle infinite-horizon tasks. For instance, to perform parkour on an infinitely long track, the policy must learn to jump, leap, pace, and transition between these modes effectively. We validate this approach in challenging bipedal control tasks: parkour and diving using a 16 DoF dynamic bipedal robot, HECTOR. Our method results in a single policy per task, solving parkour across diverse tracks and omnidirectional diving from varied heights up to 2m in simulation, showcasing versatile agility. We demonstrate successful sim-to-real transfer of parkour, including leaping over gaps up to 105 % of the leg length, jumping over blocks up to 20 % of the robot's nominal height, and pacing at speeds of up to 0.6 m/s, along with effective transitions between these modes in the real robot.

Comments:	7 pages, 6 figures
Subjects:	Robotics (cs.RO)
Cite as:	arXiv:2403.04205 [cs.RO]
	(or arXiv:2403.04205v3 [cs.RO] for this version)
	https://doi.org/10.48550/arXiv.2403.04205

Submission history

From: Lokesh Krishna [view email]
[v1] Thu, 7 Mar 2024 04:21:12 UTC (4,110 KB)
[v2] Fri, 14 Jun 2024 06:18:59 UTC (5,867 KB)
[v3] Wed, 18 Sep 2024 20:15:08 UTC (17,327 KB)

Computer Science > Robotics

Title:OGMP: Oracle Guided Multi-mode Policies for Agile and Versatile Robot Control

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Robotics

Title:OGMP: Oracle Guided Multi-mode Policies for Agile and Versatile Robot Control

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators