Topic · function approximation
function approximation
The 45 most recent episodes and tracks on this topic.
Saves to your Watch queue, to pick up on another day or another device.
Pick anything below and it plays in the bar at the foot of the window — and keeps playing while you go on browsing the directory.
- Lecture 14 - REINFORCE | Reinforcement Learning Phase|Reasoning LLMs from ScratchReasoning LLMs from ScratchNotes
- Lecture 13 - Policy Gradient Methods | Reinforcement Learning Phase | Reasoning LLMs from ScratchReasoning LLMs from ScratchNotes
- Session 21: Actor Critic based Policy Gradient, Safe RL, Planning, DYNA, Curriculum LearningJadavpur University, 2025: Introduction to Reinforcement LearningNotes
- Lecture 12 - Policy Control using Value Function Approximation | Reasoning LLMs from ScratchReasoning LLMs from ScratchNotes
- Session 20: Deep Neural Networks, MLP, Backpropagation, Policy Gradient, REINFORCEJadavpur University, 2025: Introduction to Reinforcement LearningNotes
- Session 19: Asynchronous Q learning, Classification in ML, MLE, Logistic and Softmax RegressionJadavpur University, 2025: Introduction to Reinforcement LearningNotes
- Lecture 11 - Function Approximation Methods|Reinforcement Learning Phase|Reasoning LLMs from ScratchReasoning LLMs from ScratchNotes
- Lecture 10 -Temporal Difference Control | Reinforcement Learning Phase | Reasoning LLMs from ScratchReasoning LLMs from ScratchNotes
- Session 18 Synchronous Q-learning, Model-free, based, tabular, with Linear Fn. Approx., ConvergenceJadavpur University, 2025: Introduction to Reinforcement LearningNotes
- Lecture 9 - Temporal Difference Prediction|Reinforcement Learning Phase| Reasoning LLMs from ScratchReasoning LLMs from ScratchNotes
- Session 17: Off-Policy Evaluation of TD0 with linear function Approximation, Emphatic TD0Jadavpur University, 2025: Introduction to Reinforcement LearningNotes
- Lecture 8 - Monte Carlo Methods | Reinforcement Learning Phase | Reasoning LLMs from ScratchReasoning LLMs from ScratchNotes
- Session 16 γ contraction, Banach's Fixed Point Theorem, How far is it far from the intended optimalJadavpur University, 2025: Introduction to Reinforcement LearningNotes
- Lecture 7 - Dynamic Programming | Reinforcement Learning Phase | Reasoning LLMs from ScratchReasoning LLMs from ScratchNotes
- Session 15 TD(0) convergence proof (contd), Point of Convergence of TD(0) (linear function approx.)Jadavpur University, 2025: Introduction to Reinforcement LearningNotes
- Lecture 6 - Value Functions | Reinforcement Learning | Reasoning LLMs from ScratchReasoning LLMs from ScratchNotes
- Session 14: TD0 with linear function approximation, Glimpse at Stochastic Approximation Algorithm(1)Jadavpur University, 2025: Introduction to Reinforcement LearningNotes
- Lecture 5 - Markov Decision Processes | Reasoning LLMs from ScratchReasoning LLMs from ScratchNotes
- Session 13: Function Approximation in RL, Policy Evaluation, SGD Monte Carlo, TD(0) ImplementationJadavpur University, 2025: Introduction to Reinforcement LearningNotes
- Session 12: On Policy vs Off Policy Algorithms, Importance Sampling, Model-free Q learning, SARSAJadavpur University, 2025: Introduction to Reinforcement LearningNotes
- Lecture 4b - Multi-Arm Bandits | Reasoning LLMs from ScratchReasoning LLMs from ScratchNotes
- Lecture 4 - Reinforcement Learning - Basics | Reasoning LLMs from ScratchReasoning LLMs from ScratchNotes
- Lecture 3 - Verifiers - Beam Search | Reasoning LLMs from ScratchReasoning LLMs from ScratchNotes
- Lecture 2 - Chain of Thought Reasoning | Reasoning LLMs from Scratch SeriesReasoning LLMs from ScratchNotes
- Lecture 1 - Reasoning LLMs from Scratch - Series IntroductionReasoning LLMs from ScratchNotes
- Lecture 22: Foundations of Reinforcement Learning: Partially Observable Reinforcement Learning IIPrinceton University Lectures - Foundations of Reinforcement LearningNotes
- Lecture 21: Foundations of Reinforcement Learning: Partially Observable Reinforcement Learning IPrinceton University Lectures - Foundations of Reinforcement LearningNotes
- Lecture 20: Foundations of Reinforcement Learning: Multiplayer General-Sum GamesPrinceton University Lectures - Foundations of Reinforcement LearningNotes
- Lecture 19: Foundations of Reinforcement Learning: Two-Player Zero-Sum GamesPrinceton University Lectures - Foundations of Reinforcement LearningNotes
- Lecture 17: Foundations of Reinforcement Learning: Exploration in General Function ApproximationPrinceton University Lectures - Foundations of Reinforcement LearningNotes
- Lecture 18: Foundations of Reinforcement Learning: Multiagent Reinforcement LearningPrinceton University Lectures - Foundations of Reinforcement LearningNotes
- Lecture 11: Foundations of Reinforcement Learning: Lower Bounds for MDPPrinceton University Lectures - Foundations of Reinforcement LearningNotes
- Lecture 14: Foundations of Reinforcement Learning: Least-Squares Value IterationPrinceton University Lectures - Foundations of Reinforcement LearningNotes
- Lecture 13: Foundations of Reinforcement Learning: RL in Large State SpacePrinceton University Lectures - Foundations of Reinforcement LearningNotes
- Lecture 12: Foundations of Reinforcement Learning: Offline RLPrinceton University Lectures - Foundations of Reinforcement LearningNotes
- Outlook and Research Insights (Safe, Edge and Meta Reinforcement Learning - Lecture 14, Summer 2023)Reinforcement Learning Course: Lectures (Summer 2023)Notes
- Further Contemporary RL Algorithms (TRPO, PPO - Lecture 13, Summer 2023)Reinforcement Learning Course: Lectures (Summer 2023)Notes
- Deterministic Policy Gradient Methods (Lecture 12, Summer 2023)Reinforcement Learning Course: Lectures (Summer 2023)Notes
- Stochastic Policy Gradient Methods (Lecture 11, Summer 2023)Reinforcement Learning Course: Lectures (Summer 2023)Notes
- Value-Based Control with Function Approximation (Lecture 10, Summer 2023)Reinforcement Learning Course: Lectures (Summer 2023)Notes
- On-Policy Prediction with Function Approximation (Lecture 09, Summer 2023)Reinforcement Learning Course: Lectures (Summer 2023)Notes
- Function Approximation with Supervised Learning (Lecture 08, Summer 2023)Reinforcement Learning Course: Lectures (Summer 2023)Notes
- Planning and Learning with Tabular Methods (Lecture 07, Summer 2023)Reinforcement Learning Course: Lectures (Summer 2023)Notes
- Multi-Step Bootstrapping (Lecture 06, Summer 2023)Reinforcement Learning Course: Lectures (Summer 2023)Notes
- Temporal Difference Learning (Lecture 05, Summer 2023)Reinforcement Learning Course: Lectures (Summer 2023)Notes
This playlist:.m3u.plsAll the feeds behind it
