The Inverted Pendulum Problem with Deep Reinforcement Learning | by Saif Uddin Mahmud | Dabbler in Destress | Medium
A professor of mine introduced me to the rather simple inverted pendulum problem — balance a stick on a moving platform, a hand let’s say. Intuition built on the physics of the “Game Engine” in our head tells us: if the stick is leaning to the left, move your hand to the left; if the stick is leaning to the right, move your hand to the right. We, humans, are exceptional at learning new novel tasks like these with very few sample points. Could a simple Arduino balance it? Control Theory is the obvious way to go, and having had prior experience tinkering with PID, my overconfident self walked straight into Steve Brunton’s excellent Control Bootcamp…only to return disillusioned. The task was not trivial and involved heavy math. Well if I couldn’t learn it, I’ll let the machines learn it themselves. I tried again with David Silver’s Reinforcement Learning class. Now Q-Learning and Policy Methods based on Markov Decision Processes are cool and all, but they still seemed unwieldy for continu
The Inverted Pendulum Problem with Deep Reinforcement Learning A look into Keras-RL and OpenAI libraries Saif Uddin Mahmud 10 min read · Mar 12, 2019 -- 3 Listen Share A professor of mine introduced me to the rather simple inverted pendulum problem — balance a stick on a moving platform, a hand let’s say. Intuition built on the physics of the “Game Engine” in our head tells us: if the stick is leaning to the left, move your hand to the left; if the stick is leaning to the right, move your hand to the right. We, humans, are exceptional at learning new novel tasks like these with very few sample
Explore this link on the map →saved by
related reading
- Deep Q-Networks Explained — LessWronglesswrong.com
- A (Long) Peek into Reinforcement Learning | Lil'Loglilianweng.github.io
- Debugging Reinforcement Learning Systemsandyljones.com
- An Updated Introduction to Reinforcement Learning | Sri's Blogsrianumakonda.com
- Key Papers in Deep RL - Spinning Up documentationspinningup.openai.com
- Part 1: Key Concepts in RL - Spinning Up documentationspinningup.openai.com
- Deep Reinforcement Learning: Pong from Pixelskarpathy.github.io
- Q-learning - Wikipediaen.wikipedia.org
- Deep Reinforcement Learning Doesn't Work Yetalexirpan.com
- Learn Reinforcement Learning (2) - DQN · greentec's bloggreentec.github.io
- Q-learning is not yet scalableseohong.me
- Policy Gradient Algorithms | Lil'Loglilianweng.github.io