Deep Q-Network with Experience Replay

Deep Q-Network implementation using Keras to solve the CartPole problem in OpenAI's gym. Implemented using a feed forward neural network as a non-linear value function approximator.

CartPole Problem Statement:

A pole is attached by an un-actuated joint to a cart, which moves along a frictionless track. The system is controlled by applying a force of +1 or -1 to the cart. The pendulum starts upright, and the goal is to prevent it from falling over. A reward of +1 is provided for every timestep that the pole remains upright. The episode ends when the pole is more than 15 degrees from vertical, or the cart moves more than 2.4 units from the center.

Results:

CartPole-v0 defines "solving" as getting average reward of 195.0 over 100 consecutive trials.
The DQN is very susceptible to the learning rate of the neural network.

Parameters used:

params = {
    "num_episodes": 300,
    "epsilon": 1,
    "rep_size": 20,
    "discount_factor": 0.95,
    "eps_decay": 0.995,
    "min_eps": 0.01
}

Usage:

Run pip install -r requirements.txt.
Tune hyperparameters in DQN-Code.ipynb, run the model.
Run the last cell in DQN-Code.ipynb to save the model files.
Run render.py to visualise the solution.

Name		Name	Last commit message	Last commit date
Latest commit History 8 Commits
ModelFiles		ModelFiles
images		images
.gitignore		.gitignore
DQN-Code.ipynb		DQN-Code.ipynb
README.md		README.md
render.py		render.py
requirements.txt		requirements.txt

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

Deep Q-Network with Experience Replay

CartPole Problem Statement:

Results:

Usage:

About

Releases

Packages

Contributors 2

Languages

sid-sr/Deep-Q-Network

Folders and files

Latest commit

History

Repository files navigation

Deep Q-Network with Experience Replay

CartPole Problem Statement:

Results:

Usage:

About

Topics

Resources

Stars

Watchers

Forks

Releases

Packages 0

Contributors 2

Languages

Packages