This paper employs the PPO (Proximal Policy Optimization) algorithm to study the risk hedging problem of the Shanghai Stock Exchange (SSE) 50ETF options. First, the action and state spaces were designed based on the characteristics of the hedging task, and a reward function was developed according to the cost function of the options. Second, combining the concept of curriculum learning, the agent was guided to adopt a simulated-to-real learning approach for dynamic hedging tasks, reducing the learning difficulty and addressing the issue of insufficient option data. A dynamic hedging strategy for 50ETF options was constructed. Finally, numerical experiments demonstrate the superiority of the designed algorithm over traditional hedging strategies in terms of hedging effectiveness.
Publications
- Article type
- Year
Article type
Year
Open Access
Research Article
Issue
Journal of Automation and Intelligence 2025, 4(3): 198-206
Published: 09 April 2025
Downloads:6
Total 1
京公网安备11010802044758号