RSS Amplifier

Video feed

Jadavpur University, 2025: Introduction to Reinforcement Learning

youtube.comSource feed ↗10 videos

Dormant Last read · last published · next check
Read 3 days ago and current, but nothing has been published for 15 months.

Written by

Latest videos

Saves to your Watch queue, to pick up on another day or another device.

Session 21: Actor Critic based Policy Gradient, Safe RL, Planning, DYNA, Curriculum Learning

Play

Session 20: Deep Neural Networks, MLP, Backpropagation, Policy Gradient, REINFORCE

Play

Session 19: Asynchronous Q learning, Classification in ML, MLE, Logistic and Softmax Regression

Play

Session 18 Synchronous Q-learning, Model-free, based, tabular, with Linear Fn. Approx., Convergence

Play

Session 17: Off-Policy Evaluation of TD0 with linear function Approximation, Emphatic TD0

Play

Session 16 γ contraction, Banach's Fixed Point Theorem, How far is it far from the intended optimal

Play

Session 15 TD(0) convergence proof (contd), Point of Convergence of TD(0) (linear function approx.)

Play

Session 14: TD0 with linear function approximation, Glimpse at Stochastic Approximation Algorithm(1)

Play

Session 13: Function Approximation in RL, Policy Evaluation, SGD Monte Carlo, TD(0) Implementation

Play

Session 12: On Policy vs Off Policy Algorithms, Importance Sampling, Model-free Q learning, SARSA

Play