gsurma/cartpolePublic

NotificationsYou must be signed in to change notification settings
Fork113
Star156

OpenAI's cartpole env solver.

License

MIT license

156 stars 113 forks Branches Tags Activity

Star

Notifications

You must be signed in to change notification settings

Branches Tags

Folders and files

Name		Name	Last commit message	Last commit date
Latest commit History 17 Commits
.github		.github
.idea		.idea
assets		assets
scores		scores
.gitignore		.gitignore
LICENSE		LICENSE
README.md		README.md
cartpole.py		cartpole.py
requirements.txt		requirements.txt

Repository files navigation

Cartpole

Reinforcement Learning solution of theOpenAI's Cartpole.

Check out corresponding Medium article:Cartpole - Introduction to Reinforcement Learning (DQN - Deep Q-Learning)

About

A pole is attached by an un-actuated joint to a cart, which moves along a frictionless track. The system is controlled by applying a force of +1 or -1 to the cart. The pendulum starts upright, and the goal is to prevent it from falling over. A reward of +1 is provided for every timestep that the pole remains upright. The episode ends when the pole is more than 15 degrees from vertical, or the cart moves more than 2.4 units from the center.source

DQN

Standard DQN with Experience Replay.

Hyperparameters:

GAMMA = 0.95
LEARNING_RATE = 0.001
MEMORY_SIZE = 1000000
BATCH_SIZE = 20
EXPLORATION_MAX = 1.0
EXPLORATION_MIN = 0.01
EXPLORATION_DECAY = 0.995

Model structure:

Dense layer - input:4, output:24, activation:relu
Dense layer - input24, output:24, activation:relu
Dense layer - input24, output:2, activation:linear

MSE loss function
Adam optimizer

Performance

CartPole-v0 defines "solving" as getting average reward of 195.0 over 100 consecutive trials.source

Example trial gif

Example trial chart

Solved trials chart

Author

Greg (Grzegorz) Surma

PORTFOLIO

GITHUB

BLOG

About

OpenAI's cartpole env solver.

gsurma.github.io

Releases

No releases published

Sponsor this project

patreon.com/gsurma

Packages

No packages published

Contributors2

Languages

Python100.0%

Movatterモバイル変換

Navigation Menu

Search code, repositories, users, issues, pull requests...

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Uh oh!

License

Uh oh!

Folders and files

Latest commit

History

Repository files navigation

Cartpole

About

DQN

Hyperparameters:

Model structure:

Performance

Example trial gif

Example trial chart

Solved trials chart

Author

About

Topics

Resources

License

Uh oh!

Stars

Watchers

Forks

Releases

Sponsor this project

Uh oh!

Packages

Uh oh!

Contributors2

Uh oh!

Languages

Movatterモバイル変換

Uh oh!

License

gsurma/cartpole

Folders and files

Latest commit

History

Repository files navigation

Cartpole

About

DQN

Hyperparameters:

Model structure:

Performance

Example trial gif

Example trial chart

Solved trials chart

Author

About

Topics

Resources

License

Uh oh!

Stars

Watchers

Forks

Releases

Sponsor this project

Uh oh!

Packages0

Uh oh!

Contributors2

Uh oh!

Languages

Packages