AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
Article Link
Collect
Submit Manuscript
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Full Length Article | Open Access

The investigation of reinforcement learning-based end-to-end decision-making algorithms for autonomous driving on the road with consecutive sharp turns

Tongyang LiJiageng Ruan( )Kaixuan Zhang
College of Mechanical and Energy Engineering, Beijing University of Technology, Beijing 100020, China
Show Author Information

HIGHLIGHTS

• Three deep reinforcement learning-based decision-making policies are proposed.

• The role of observation variable in agent training quality is analyzed.

• A novel reward setting method is proposed to solve the sparse reward problem.

• Comprehensive comparisons are made between strategies for consecutive sharp turns.

Abstract

Learning-based algorithm attracts great attention in the autonomous driving control field, especially for decision-making, to meet the challenge in long-tail extreme scenarios, where traditional methods demonstrate poor adaptability even with a significant effort. To improve the autonomous driving performance in extreme scenarios, specifically consecutive sharp turns, three deep reinforcement learning algorithms, i.e. Deep Deterministic Policy Gradient (DDPG), Twin Delayed Deep Deterministic policy gradient (TD3), and Soft Actor-Critic (SAC), based decision-making policies are proposed in this study. The role of the observation variable in agent training is discussed by comparing the driving stability, average speed, and consumed computational effort of the proposed algorithms in curves with various curvatures. In addition, a novel reward-setting method that combines the states of the environment and the vehicle is proposed to solve the sparse reward problem in the reward-guided algorithm. Simulation results from the road with consecutive sharp turns show that the DDPG, SAC, and TD3 algorithms-based vehicles take 367.2, 359.6, and 302.1 ​s to finish the task, respectively, which match the training results, and verifies the observation variable role in agent quality improvement.

Graphical Abstract

Electronic Supplementary Material

Download File(s)
geits-4-3-100288_ESM.pdf (54.1 KB)

References

【1】
【1】
 
 
Green Energy and Intelligent Transportation

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Li T, Ruan J, Zhang K. The investigation of reinforcement learning-based end-to-end decision-making algorithms for autonomous driving on the road with consecutive sharp turns. Green Energy and Intelligent Transportation, 2025, 4(3). https://doi.org/10.1016/j.geits.2025.100288

1141

Views

45

Crossref

45

Web of Science

55

Scopus

Received: 24 May 2024
Revised: 12 December 2024
Accepted: 30 December 2024
Published: 18 March 2025
© 2025

This is an open access article under the CC BY-NC-ND license (http://creativecommons.org/licenses/by-nc-nd/4.0/).