AI04, Contents

AI04 : Reinforcement learning

The Markov Decision Process and Dynamic Programming
Gaming with Monte Carlo Methods
Temporal Difference Learning
Multi-Armed Bandit Problem
Deep Learning Fundamentals
Atari Games with Deep Q Network
Playing Doom with a Deep Recurrent Q Network
The Asynchronous Advantage Actor Critic Network
Policy Gradients and Optimization
Car Racing Using DQN
Recent Advancements and Next Steps