Engineering PapersSearch

DOE OSTI · code-140520

PPO And Friends

Abstract

PPO and Friends (PPOAF) is a pytorch implementation of proximal policy optimization for single- and multi-agent reinforcement learning (the PPO), along with several optimizations and add-ons (the Friends) to enable efficient MPI-parallelized model training on HPC clusters.

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Maguire, AlisterO. 2024-05-31. PPO And Friends. https://doi.org/10.11578/dc.20240815.4

Cite the original work for its findings. Save a collection to share your selection of sources.