Normalized Advantage Functions (NAF) in TensorFlow

TensorFlow implementation of Continuous Deep q-Learning with Model-based Acceleration.

Requirements

Python 2.7
gym
TensorFlow 0.9+

Usage

First, install prerequisites with:

$ pip install tqdm gym[all]

To train a model for an environment with a continuous action space:

$ python main.py --env_name=Pendulum-v0 --is_train=True
$ python main.py --env_name=Pendulum-v0 --is_train=True --display=True

To test and record the screens with gym:

$ python main.py --env_name=Pendulum-v0 --is_train=False
$ python main.py --env_name=Pendulum-v0 --is_train=False --display=True

Results

Training details of Pendulum-v0 with different hyperparameters.

$ python main.py --env_name=Pendulum-v0 # dark green
$ python main.py --env_name=Pendulum-v0 --action_fn=tanh # light green
$ python main.py --env_name=Pendulum-v0 --use_batch_norm=True # yellow
$ python main.py --env_name=Pendulum-v0 --use_seperate_networks=True # green

References

Author

Taehoon Kim / @carpedm20

Name		Name	Last commit message	Last commit date
Latest commit History 42 Commits
assets		assets
src		src
.gitignore		.gitignore
LICENSE		LICENSE
README.md		README.md
main.py		main.py
run_mujoco.sh		run_mujoco.sh
utils.py		utils.py

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

assets

assets

src

src

.gitignore

.gitignore

LICENSE

LICENSE

README.md

README.md

main.py

main.py

run_mujoco.sh

run_mujoco.sh

utils.py

utils.py

Repository files navigation

Normalized Advantage Functions (NAF) in TensorFlow

Requirements

Usage

Results

References

Author

About

Releases

Packages

Contributors 2

Languages

License

carpedm20/NAF-tensorflow

Folders and files

Latest commit

History

Repository files navigation

Normalized Advantage Functions (NAF) in TensorFlow

Requirements

Usage

Results

References

Author

About

Topics

Resources

License

Stars

Watchers

Forks

Languages