AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
Article Link
Collect
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Open Access

Continuous-time hierarchical reinforcement learning for satellite pursuit decision

Linsen WEIaXin NINGbXiaobin LIANb( )Feng WANGcGaopeng ZHANGcMingpei LINd
National Elite Institute of Engineering, Northwestern Polytechnical University, Xi’an 710072, China
School of Astronautics, Northwestern Polytechnical University, Xi’an 710072, China
Xi’an Institute of Optics and Precision Mechanics, Chinese Academy of Sciences, Xi’an 710119, China
Institute of Multidisciplinary Research for Advanced Materials, Tohoku University, Sendai 980-8577, Japan

Peer review under responsibility of Editorial Committee of CJA.

Show Author Information

Abstract

The satellite orbital pursuit game focuses on studying spacecraft maneuvering strategies in space. Traditional numerical methods often face real-time inadequacies and adaptability limitations when dealing with highly nonlinear problems. With the advancement of Deep Reinforcement Learning (DRL) technology, continuous-time orbital control capabilities have significantly improved. Despite this, the existing DRL technologies still need adjustments in action delay and discretization structure to better adapt to practical application scenarios. Combining continuous learning and model planning demonstrates the adaptability of these methods in continuous-time decision problems. Additionally, to more effectively handle action delay issues, a new scheduled action execution technique has been developed. This technique optimizes action execution timing through real-time policy adjustments, thus adapting to the dynamic changes in the orbital environment. A Hierarchical Reinforcement Learning (HRL) strategy was also adopted to simplify the decision-making process for long-distance pursuit tasks by setting phased subgoals to gradually approach the target. The effectiveness of the proposed strategy in practical satellite pursuit scenarios has been verified through simulations of two different tasks.

References

【1】
【1】
 
 
Chinese Journal of Aeronautics

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
WEI L, NING X, LIAN X, et al. Continuous-time hierarchical reinforcement learning for satellite pursuit decision. Chinese Journal of Aeronautics, 2025, 38(12). https://doi.org/10.1016/j.cja.2025.103662

624

Views

2

Crossref

2

Web of Science

2

Scopus

0

CSCD

Received: 11 October 2024
Revised: 13 November 2024
Accepted: 06 December 2024
Published: 08 July 2025
© 2025 The Authors. Chinese Society of Aeronautics and Astronautics.

This is an open access article under the CC BY-NC-ND license (http://creativecommons.org/licenses/by-nc-nd/4.0/).