AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
Article Link
Collect
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Article | Open Access

Dynamic path planning of autonomous mobile robot in off-road environments using experience replay enhanced distributed proximal policy optimization algorithm

Qingyun Liua Xiong Youa ( )Xin Zhanga Jiwei Zuob Jia Lic 
Institute of Geographical Spatial Information, Information Engineering University, Zhengzhou, China
School of Civil Engineering and Geomatics, Shandong University of Technology, Zibo, China
Institute of Data and Target Engineering, Information Engineering University, Zhengzhou, China
Show Author Information

Abstract

In off-road environments, some problems like insufficient consideration of environmental trafficability, difficulty in model convergence, and limited generalization often exist in autonomous mobile robot (AMR) path planning research based on reinforcement learning (RL). In order to solve these problems, a distributed proximal policy optimization dynamic path planning algorithm based on experience playback enhancement (Re-DPPO) is proposed, which realizes the safe, feasible, and optimal trafficability path planning of AMR between any starting and ending points in an off-road environment. First, in order to achieve AMR path planning that takes into account environmental trafficability in off-road environments, an AMR trafficability map that integrates multiple environmental factors was constructed. Second, to ensure that AMR can simultaneously consider path trafficability during real-time obstacle avoidance, a multi-dimensional comprehensive reward function was designed that integrates dynamic obstacle avoidance, goal point proximity, and trafficability evaluation. In addition, in response to the problems of difficult model convergence and low utilization of high-value samples in traditional deep reinforcement learning (DRL) algorithms, a distributed training architecture and a priority experience replay mechanism are introduced. Finally, comparative experiments were conducted between the Re-DPPO algorithm and five DRL algorithms, namely the proximal policy optimization (PPO), experience replay enhanced proximal policy optimization algorithm (Re-PPO), distributed proximal policy optimization (DPPO), deep Q-network (DQN), and double DQN, and the path planning test was carried out in four new scenarios. The results show that among the six algorithms, the Re-DPPO algorithm exhibits the best convergence performance and path planning success rate. This algorithm can prioritize planning the optimal trafficability path while ensuring accessibility, and demonstrates good generalization ability in all four new scenarios.

References

【1】
【1】
 
 
Geo-Spatial Information Science
Pages 3120-3136

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Liu Q, You X, Zhang X, et al. Dynamic path planning of autonomous mobile robot in off-road environments using experience replay enhanced distributed proximal policy optimization algorithm. Geo-Spatial Information Science, 2026, 29(4): 3120-3136. https://doi.org/10.1080/10095020.2026.2624861

3

Views

0

Crossref

0

Web of Science

0

Scopus

0

CSCD

Received: 03 June 2025
Accepted: 26 January 2026
Published: 16 February 2026
© 2026 Wuhan University.

This is an Open Access article distributed under the terms of the Creative Commons Attribution License (http://creativecommons.org/licenses/by/4.0/), which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited. The terms on which this article has been published allow the posting of the Accepted Manuscript in a repository by the author(s) or with their consent.