Engineering PapersSearch

NASA NTRS · 20230005674

Multi-objective Reinforcement Learning for Low-thrust Transfer Design Between Libration Point Orbits

Abstract

Multi-Reward Proximal Policy Optimization (MRPPO) is a multi-objective rein- forcement learning algorithm used to construct low-thrust transfers between pe- riodic orbits in multi-body systems. Previous implementations of MRPPO have relied on a predefined reference transfer to successfully train each policy. In this paper, an algorithmic modification labeled the ‘moving reference’, is introduced to autonomously construct these reference trajectories during training. With this modification, MRPPO is used to recover various low-thrust transfers between two periodic orbits in the Earth-Moon circular restricted three-body problem to solve a multi-objective optimization problem. These results are then compared with the solutions recovered via a gradient descent optimization scheme to validate the performance of MRPPO with the moving reference modification.

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Anderson, Rodney L, Mashiku, Alinda K., Bosanac, Natasha, Sullivan, Christopher J.. 2021-08-09. Multi-objective Reinforcement Learning for Low-thrust Transfer Design Between Libration Point Orbits. https://ntrs.nasa.gov/citations/20230005674

Cite the original work for its findings. Save a collection to share your selection of sources.