Learning Risk-Aware Quadrupedal Locomotion using Distributional Reinforcement Learning

Schneider, Lukas; Frey, Jonas; Miki, Takahiro; Hutter, Marco

Computer Science > Robotics

arXiv:2309.14246 (cs)

[Submitted on 25 Sep 2023 (v1), last revised 3 May 2024 (this version, v2)]

Title:Learning Risk-Aware Quadrupedal Locomotion using Distributional Reinforcement Learning

Authors:Lukas Schneider, Jonas Frey, Takahiro Miki, Marco Hutter

View PDF HTML (experimental)

Abstract:Deployment in hazardous environments requires robots to understand the risks associated with their actions and movements to prevent accidents. Despite its importance, these risks are not explicitly modeled by currently deployed locomotion controllers for legged robots. In this work, we propose a risk sensitive locomotion training method employing distributional reinforcement learning to consider safety explicitly. Instead of relying on a value expectation, we estimate the complete value distribution to account for uncertainty in the robot's interaction with the environment. The value distribution is consumed by a risk metric to extract risk sensitive value estimates. These are integrated into Proximal Policy Optimization (PPO) to derive our method, Distributional Proximal Policy Optimization (DPPO). The risk preference, ranging from risk-averse to risk-seeking, can be controlled by a single parameter, which enables to adjust the robot's behavior dynamically. Importantly, our approach removes the need for additional reward function tuning to achieve risk sensitivity. We show emergent risk sensitive locomotion behavior in simulation and on the quadrupedal robot ANYmal. Videos of the experiments and code are available at this https URL.

Subjects:	Robotics (cs.RO); Machine Learning (cs.LG)
Cite as:	arXiv:2309.14246 [cs.RO]
	(or arXiv:2309.14246v2 [cs.RO] for this version)
	https://doi.org/10.48550/arXiv.2309.14246

Submission history

From: Lukas Schneider [view email]
[v1] Mon, 25 Sep 2023 16:05:32 UTC (37,142 KB)
[v2] Fri, 3 May 2024 04:39:46 UTC (37,155 KB)

Computer Science > Robotics

Title:Learning Risk-Aware Quadrupedal Locomotion using Distributional Reinforcement Learning

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Robotics

Title:Learning Risk-Aware Quadrupedal Locomotion using Distributional Reinforcement Learning

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators