AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (5.5 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Open Access

Reset-Free Reinforcement Learning via Multi-State Recovery and Failure Prevention for Autonomous Robots

Xu Zhou1,2Benlian Xu3( )Zhengqiang Jiang4Jun Li2Brett Nener5
School of Mechanical Engineering, Changshu Institute of Technology, Changshu 215500, China
School of Automation, Nanjing University of Science and Technology, Nanjing 210094, China
School of Electronic and Information Engineering, Suzhou University of Science and Technology, Suzhou 215009, China
Faculty of Medicine and Health, The University of Sydney, Sydney 2006, Australia
Department of Electrical, Electronic and Computer Engineering, The University of Western Australia, Perth 6009, Australia
Show Author Information

Abstract

Reinforcement learning holds promise in enabling robotic tasks as it can learn optimal policies via trial and error. However, the practical deployment of reinforcement learning usually requires human intervention to provide episodic resets when a failure occurs. Since manual resets are generally unavailable in autonomous robots, we propose a reset-free reinforcement learning algorithm based on multi-state recovery and failure prevention to avoid failure-induced resets. The multi-state recovery provides robots with the capability of recovering from failures by self-correcting its behavior in the problematic state and, more importantly, deciding which previous state is the best to return to for efficient re-learning. The failure prevention reduces potential failures by predicting and excluding possible unsafe actions in specific states. Both simulations and real-world experiments are used to validate our algorithm with the results showing a significant reduction in the number of resets and failures during the learning.

References

【1】
【1】
 
 
Tsinghua Science and Technology
Pages 1481-1494

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Zhou X, Xu B, Jiang Z, et al. Reset-Free Reinforcement Learning via Multi-State Recovery and Failure Prevention for Autonomous Robots. Tsinghua Science and Technology, 2024, 29(5): 1481-1494. https://doi.org/10.26599/TST.2023.9010117
Part of a topical collection:

2200

Views

108

Downloads

3

Crossref

3

Web of Science

3

Scopus

0

CSCD

Received: 13 August 2023
Revised: 02 October 2023
Accepted: 10 October 2023
Published: 02 May 2024
© The Author(s) 2024.

The articles published in this open access journal are distributed under the terms of the Creative Commons Attribution 4.0 International License (http://creativecommons.org/licenses/by/4.0/).