
Mathematical Foundations of Reinforcement Learning
Dormant Last read · last published · next check
Read 4 days ago and current, but nothing has been published for 24 months.
Latest videos
Saves to your Watch queue, to pick up on another day or another device.


L4: Value Iteration and Policy Iteration (P2-Policy iteration)—Mathematical Foundations of RL

L4: Value Iteration and Policy Iteration (P1-Value iteration)—Mathematical Foundations of RL

L3: Bellman Optimality Equation (P4-Interesting properties)—Mathematical Foundations of RL

L3: Bellman Optimality Equation (P3-More)—Mathematical Foundations of RL

L3: Bellman Optimality Equation (P2-Optimal policy)—Mathematical Foundations of RL

L3: Bellman Optimality Equation (P1-Motivating example)—Mathematical Foundations of RL

L2: Bellman Equation (P4-Matrix-vector form and solution)—Mathematical Foundations of RL

L2: Bellman Equation (P5-Action value)—Mathematical Foundations of RL

