Courselet on Reinforcement Learning, giving an entrance level description on problem formulation as Markov Decision Processes, the three curses of dimensionality and basic temporal difference methods.
Courselet on Reinforcement Learning, giving an entrance level description on problem formulation as Markov Decision Processes, the three curses of dimensionality and basic temporal difference methods. The courselet treats the following topics:
- Building blocks of Markov Decision Processes
- State-, action- and outcome spaces
- Post-decision states and look-up tables
- Introduction to temporal difference learning: Monte carlo learning, Q-learning and SARSA
WP6 (Doctoral Training) coordinator and teacher in MSCA DIGITAL