Name		Name	Last commit message	Last commit date
parent directory ..
Images		Images
model		model
runs		runs
Categorical_DQN.py		Categorical_DQN.py
LICENSE		LICENSE
README.md		README.md
main.py		main.py
utils.py		utils.py

README.md

C51: Categorical-DQN-Pytorch

A clean and robust Pytorch implementation of Categorical DQN(C51) ：

Render	Training curve

Other RL algorithms by Pytorch can be found here.

Dependencies

gymnasium==0.29.1
matplotlib==3.8.2
numpy==1.26.1
pytorch==2.1.0

python==3.11.5

How to use my code

Train from scratch

python main.py

where the default enviroment is 'CartPole'.

Change Enviroment

If you want to train on different enviroments, just run:

python main.py --EnvIdex 1

The --EnvIdex can be set to be 0 and 1, where

'--EnvIdex 0' for 'CartPole-v1'  
'--EnvIdex 1' for 'LunarLander-v2'

Note: if you want to play on LunarLander, you need to install box2d-py first. You can install box2d-py via: pip install gymnasium[box2d]

Play with trained model

python main.py --EnvIdex 0 --render True --Loadmodel True --ModelIdex 60 # Play with CartPole

python main.py --EnvIdex 1 --render True --Loadmodel True --ModelIdex 320 # Play with LunarLander

Visualize the training curve

You can use the tensorboard to record anv visualize the training curve.

Installation (please make sure Pytorch is installed already):

pip install tensorboard
pip install packaging

Record (the training curves will be saved at '\runs'):

python main.py --write True

Visualization:

tensorboard --logdir runs

Hyperparameter Setting

For more details of Hyperparameter Setting, please check 'main.py'

References

Bellemare M G, Dabney W, Munos R. A distributional perspective on reinforcement learning[C]//International conference on machine learning. PMLR, 2017: 449-458.

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

2.4_Categorical-DQN_C51

2.4_Categorical-DQN_C51

README.md

C51: Categorical-DQN-Pytorch

Dependencies

How to use my code

Train from scratch

Change Enviroment

Play with trained model

Visualize the training curve

Hyperparameter Setting

References

Files

2.4_Categorical-DQN_C51

Directory actions

More options

Directory actions

More options

Latest commit

History

2.4_Categorical-DQN_C51

Folders and files

parent directory

README.md

C51: Categorical-DQN-Pytorch

Dependencies

How to use my code

Train from scratch

Change Enviroment

Play with trained model

Visualize the training curve

Hyperparameter Setting

References