| ★ | An Introduction to Reinforcement Learning | | 247 | Estados Unidos | |
| 11 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Tutorial Session: Review of Q-Learning | Stanford Online @stanfordonline | — | Estados Unidos | 86 |
| 16 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 16: RL for Robots | Stanford Online @stanfordonline | — | Estados Unidos | 86 |
| 19 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 1: Class Intro | Stanford Online @stanfordonline | 84.0K | Estados Unidos | 86 |
| 17 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 17: Advancing Robot Intelligence | Stanford Online @stanfordonline | 5.5K | Estados Unidos | 86 |
| 13 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 11: Model-Based RL | Stanford Online @stanfordonline | 3.2K | Estados Unidos | 86 |
| 18 | Options (RL) expliqué : Maîtriser la planification hiérarchique | Deep Learner, One Step at a Time @deeplearneronestepatatime | 2.0K | Francia | 86 |
| 9 | Reinforcement Learning: Zero to Hero | CodeEmporium @codeemporium | 1.4K | Estados Unidos | 87 |
| 7 | Temporal Difference Learning | | 670 | — | 87 |
| 2 | What is Reinforcement Learning? | | 581 | — | 90 |
| 5 | Markov Decision Process (MDP) | | 543 | — | 88 |
| 10 | MASTERING MACHINES: The Reinforcement Learning Revolution in AI | TechWaveWeekly @techwaveweekly-zg9jr | 406 | Austria | 87 |
| 20 | Q-learning with Flow-Matching Policies | Microsoft Research @microsoftresearch | 363 | Estados Unidos | 86 |
| 6 | Q-Learning | | 259 | — | 88 |
| 3 | Lecture 5.2 | Reinforcement learning models | Q Learning | Markov Decision Process | #mlt #aktu #ml | Tech Master Edu @techmasteredu | 250 | India | 89 |
| 12 | Actor-Critic Methods | | 170 | — | 86 |
| 8 | Exploration vs Exploitation in AI and machine learning | | 120 | — | 87 |
| 14 | Policy Gradient Methods | | 113 | — | 86 |
| 15 | Machine Learning Techniques for Agents (10 Minutes) | Technology Whisper @technologywhisper | 102 | Singapur | 86 |
| 4 | Lecture 5.1 | Reinforcement learning | Reinforcement learning in practice | #mlt #aktu #unit5 | Tech Master Edu @techmasteredu | 67 | India | 88 |
| 1 | Reinforcement Learning: How Machines Learn Through Rewards | 머니렙아크 MoneyRepArc @머니렙아크moneyreparc | 0 | Corea del Sur | 90 |