NASA NTRS · 20240000230
Environment Adversarial Reinforcement Learning
Abstract
This paper presents a training method for increasing performance of reinforcement learning agents. The method is named Environment Adversarial Reinforcement Learning. The method requires the reinforcement learning environment to be parameterizeable. Over the course of training, environment parameters are updated in a direction of increasing difficulty for the agent. The direction for these updates is found using a performance prediction network trained on data from tests of the agent under varying environment parameters. The method was tested on a CartPole environment. A 28-58\% improvement in mean return was found when comparing performance to a baseline reinforcement learning algorithm on both easy and hard versions of the task.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
John R Cooper. Environment Adversarial Reinforcement Learning. https://ntrs.nasa.gov/citations/20240000230
Cite the original work for its findings. Save a collection to share your selection of sources.