RSS Amplifier

Video feed

EE675 (2024) Introduction to Reinforcement Learning Course | IIT Kanpur

youtube.comSource feed ↗10 videos

Dormant Last read · last published · next check
Read 2 days ago and current, but nothing has been published for 2 years.

Written by

Latest videos

Saves to your Watch queue, to pick up on another day or another device.

Lecture 2 | Multi-arm Bandits | Reinforcement Learning Course | IIT Kanpur

Play

Lecture 1 | Machine Learning Paradigms - An Overview | Reinforcement Learning Course | IIT Kanpur

Play

Lecture 13 | Value Iteration and Monte Carlo Prediction | Reinforcement Learning Course | IIT Kanpur

Play

Lecture 10 - Bellman Expectation Equations for MDP | Reinforcement Learning Course | IIT Kanpur

Play

Lecture 11 | Bellman Optimality Eqs | Policy Iteration | Reinforcement Learning Course | IIT Kanpur

Play

Lecture 12 | Convergence Proof of Policy Iteration | Reinforcement Learning Course | IIT Kanpur

Play

Lecture 4 - Regret Analysis of UCB Bandit algorithm | Reinforcement Learning | IIT Kanpur

Play

Lecture 5 - UCB Regret | KL Divergence | B–H inequality | Reinforcement Learning Course | IIT Kanpur

Play

Lecture 6 - Lower Bound on Regret for Bandit Algorithms | Reinforcement Learning Course | IIT Kanpur

Play

Lecture 7 - Thompson Sampling for Multi-arm Bandits | Reinforcement Learning Course | IIT Kanpur

Play