Clean, Robust, and Unified PyTorch implementation of popular Deep Reinforcement Learning (DRL) algorithms (Q-learning, Duel DDQN, PER, C51, Noisy DQN, PPO, DDPG, TD3, SAC, ASL)
Python
3,448
115 commits
updated Jun 11, 2025
This repository uses the following python dependencies unless explicitly stated:
gymnasium==0.29.1
numpy==1.26.1
pytorch==2.1.0
python==3.11.5
Enter the folder of the algorithm that you want to use, and run the main.py to train from scratch:
python main.py
For more details, please check the README.md file in the corresponding algorithm folder.
NoisyNet DQN: Fortunato M, Azar M G, Piot B, et al. Noisy networks for exploration[J]. arXiv preprint arXiv:1706.10295, 2017.
ColorDynamic: Generalizable, Scalable, Real-time, End-to-end Local Planner for Unstructured and Dynamic Environments
@misc{DRL-Pytorch,
author = {Jinghao Xin},
title = {DRL-Pytorch},
year = {2022},
publisher = {GitHub},
journal = {GitHub Repository},
howpublished = {\url{https://github.com/XinJingHao/DRL-Pytorch}},
}
| CartPole | LunarLander |
|---|---|
![]() | ![]() |
| Pong | Enduro |
|---|---|
![]() | ![]() |
| CartPole | LunarLander |
|---|---|
| CartPole | LunarLander |
|---|---|
| CartPole | LunarLander |
|---|---|
![]() | ![]() |
| Pendulum | LunarLanderContinuous |
|---|---|
Clean, Robust, and Unified PyTorch implementation of popular Deep Reinforcement Learning (DRL) algorithms (Q-learning, Duel DDQN, PER, C51, Noisy DQN, PPO, DDPG, TD3, SAC, ASL)
Python
3,448
115 commits
updated Jun 11, 2025
This repository uses the following python dependencies unless explicitly stated:
gymnasium==0.29.1
numpy==1.26.1
pytorch==2.1.0
python==3.11.5
Enter the folder of the algorithm that you want to use, and run the main.py to train from scratch:
python main.py
For more details, please check the README.md file in the corresponding algorithm folder.
NoisyNet DQN: Fortunato M, Azar M G, Piot B, et al. Noisy networks for exploration[J]. arXiv preprint arXiv:1706.10295, 2017.
ColorDynamic: Generalizable, Scalable, Real-time, End-to-end Local Planner for Unstructured and Dynamic Environments
@misc{DRL-Pytorch,
author = {Jinghao Xin},
title = {DRL-Pytorch},
year = {2022},
publisher = {GitHub},
journal = {GitHub Repository},
howpublished = {\url{https://github.com/XinJingHao/DRL-Pytorch}},
}
| CartPole | LunarLander |
|---|---|
![]() | ![]() |
| Pong | Enduro |
|---|---|
![]() | ![]() |
| CartPole | LunarLander |
|---|---|
| CartPole | LunarLander |
|---|---|
| CartPole | LunarLander |
|---|---|
![]() | ![]() |
| Pendulum | LunarLanderContinuous |
|---|---|