AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
Article Link
Collect
Submit Manuscript
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Research Article

Learning-efficient orbit decision via reinforcement learning with astrodynamics-informed action parameterization

Linsen Wei1, Changyou Li2, Xin Ning2, Yu Jiang3, Xiaobin Lian2( )
National Elite Institute of Engineering, Northwestern Polytechnical University, Xi’an 710072, China
School of Astronautics, Northwestern Polytechnical University, Xi’an 710072, China
Space Traffic Research Center, National Key Laboratory of Space System Operation and Control, Xi’an 710049, China
Show Author Information

Abstract

Satellite orbit decision-making requires a reliable approach in uncertain environments. Classical guidance methods can produce high-quality trajectories but rely on repeated numerical solutions of boundary-value problems, which limits their scalability for large-scale online decision-making. Deep reinforcement learning (RL) can in principle handle nonlinear high-dimensional dynamics, yet its interaction demand makes sample efficiency a central bottleneck for orbital decision-making under high-fidelity astrodynamics models. This paper proposes an RL framework for learning-efficient orbit decision based on an astrodynamics-informed parameterized action space. Each action consists of a discrete maneuver index and a continuous parameter vector. The discrete index selects a maneuver primitive from a library of classical orbital guidance methods. The continuous parameters unify the internal degrees of freedom of these primitives. This architecture with reparameterized continuous sampling and projection enforces maneuver feasibility and generates bounded physically meaningful commands. Comprehensive simulations demonstrate that the proposed astrodynamics-informed parameterization improves sample efficiency and final performance compared with unstructured Δ v action spaces and normalized baselines.

Graphical Abstract

References

【1】
【1】
 
 
Astrodynamics
Pages 883-900

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Wei L, Li C, Ning X, et al. Learning-efficient orbit decision via reinforcement learning with astrodynamics-informed action parameterization. Astrodynamics, 2026, 10(5): 883-900. https://doi.org/10.1007/s42064-026-0313-9

8

Views

0

Crossref

0

Web of Science

0

Scopus

0

CSCD

Received: 28 November 2025
Accepted: 09 April 2026
Published: 10 October 2026
© Tsinghua University Press 2026