AI04, Contents
AI04 : Reinforcement learning
The Markov Decision Process and Dynamic Programming
Gaming with Monte Carlo Methods
Temporal Difference Learning
Multi-Armed Bandit Problem
Deep Learning Fundamentals
Atari Games with Deep Q Network
Playing Doom with a Deep Recurrent Q Network
The Asynchronous Advantage Actor Critic Network
Policy Gradients and Optimization
Car Racing Using DQN
Recent Advancements and Next Steps