DOE OSTI · code-140520
PPO And Friends
Abstract
PPO and Friends (PPOAF) is a pytorch implementation of proximal policy optimization for single- and multi-agent reinforcement learning (the PPO), along with several optimizations and add-ons (the Friends) to enable efficient MPI-parallelized model training on HPC clusters.
Keep this discovery
Explore connections, maps & timelines
Maguire, AlisterO. 2024-05-31. PPO And Friends. https://doi.org/10.11578/dc.20240815.4
Cite the original work for its findings. Save a collection to share your selection of sources.