[ Web Proxy ]
URL:
Viewing: https://docs.cleanrl.dev/ [Back]  [Original]

CleanRL
Skip to content
CleanRL
Overview
[Go]
vwxyzjn/cleanrl
Table of contents

CleanRL - Overview

license-MIT-blue [license-MIT-blue] tests [tests] docs [docs] 767863440248143916 [767863440248143916] UCDdC6BIFRI0jvcwuhi3aI6w [UCDdC6BIFRI0jvcwuhi3aI6w] Code style: black [Code style: black] Imports: isort [Imports: isort] %F0%9F%A4%97%20Models-Huggingface-F8D521 [%F0%9F%A4%97%20Models-Huggingface-F8D521] Open In Colab [Open In Colab]

CleanRL is a Deep Reinforcement Learning library that provides high-quality single-file implementation with research-friendly features. The implementation is clean and simple, yet we can scale it to run thousands of experiments using AWS Batch. The highlight features of CleanRL are:

  • Single-file implementation
  • Every detail about an algorithm variant is put into a single standalone file.
  • For example, our ppo_atari.py only has 340 lines of code but contains all implementation details on how PPO works with Atari games, so it is a great reference implementation to read for folks who do not wish to read an entire modular library.
  • Benchmarked Implementation (7+ algorithms and 34+ games at https://benchmark.cleanrl.dev)
  • Tensorboard Logging
  • Local Reproducibility via Seeding
  • Videos of Gameplay Capturing
  • Experiment Management with Weights and Biases
  • Cloud Integration with docker and AWS

You can read more about CleanRL in our technical paper and documentation.

CleanRL only contains implementations of online deep reinforcement learning algorithms. If you are looking for offline algorithms, please check out corl-team/CORL, which shares a similar design philosophy as CleanRL.

Info

Support for Gymnasium: Farama-Foundation/Gymnasium is the next generation of openai/gym that will continue to be maintained and introduce new features. Please see their announcement for further detail. We are migrating to gymnasium and the progress can be tracked in vwxyzjn/cleanrl#277.

Warning

CleanRL is not a modular library and therefore it is not meant to be imported. At the cost of duplicate code, we make all implementation details of a DRL algorithm variant easy to understand, so CleanRL comes with its own pros and cons. You should consider using CleanRL if you want to 1) understand all implementation details of an algorithm's variant or 2) prototype advanced features that other modular DRL libraries do not support (CleanRL has minimal lines of code so it gives you great debugging experience and you don't have do a lot of subclassing like sometimes in modular DRL libraries).

Citing CleanRL

If you use CleanRL in your work, please cite our technical paper:

@article{huang2022cleanrl,
  author  = {Shengyi Huang and Rousslan Fernand Julien Dossa and Chang Ye and Jeff Braga and Dipam Chakraborty and Kinal Mehta and Joo G.M. Arajo},
  title   = {CleanRL: High-quality Single-file Implementations of Deep Reinforcement Learning Algorithms},
  journal = {Journal of Machine Learning Research},
  year    = {2022},
  volume  = {23},
  number  = {274},
  pages   = {1--18},
  url     = {http://jmlr.org/papers/v23/21-1342.html}
}
Back to top

Web Proxy Viewer  |  New URL  |  Original Page