
Jadavpur University, 2025: Introduction to Reinforcement Learning
Dormant Last read · last published · next check
Read 3 days ago and current, but nothing has been published for 15 months.
Latest videos
Saves to your Watch queue, to pick up on another day or another device.


Session 20: Deep Neural Networks, MLP, Backpropagation, Policy Gradient, REINFORCE

Session 19: Asynchronous Q learning, Classification in ML, MLE, Logistic and Softmax Regression

Session 18 Synchronous Q-learning, Model-free, based, tabular, with Linear Fn. Approx., Convergence

Session 17: Off-Policy Evaluation of TD0 with linear function Approximation, Emphatic TD0

Session 16 γ contraction, Banach's Fixed Point Theorem, How far is it far from the intended optimal

Session 15 TD(0) convergence proof (contd), Point of Convergence of TD(0) (linear function approx.)

Session 14: TD0 with linear function approximation, Glimpse at Stochastic Approximation Algorithm(1)

Session 13: Function Approximation in RL, Policy Evaluation, SGD Monte Carlo, TD(0) Implementation

