| ★ | Lecture 5.2 | Reinforcement learning models | Q Learning | Markov Decision Process | #mlt #aktu #ml | Tech Master Edu @techmasteredu | 250 | India | |
| 2 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Tutorial Session: Review of Q-Learning | Stanford Online @stanfordonline | — | Estados Unidos | 91 |
| 3 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 6: Q-Learning | Stanford Online @stanfordonline | — | Estados Unidos | 90 |
| 7 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 15: Hierarchical RL and IL | Stanford Online @stanfordonline | — | Estados Unidos | 90 |
| 10 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 18: Frontiers | Stanford Online @stanfordonline | — | Estados Unidos | 89 |
| 18 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 9: RL for LLMs | Stanford Online @stanfordonline | — | Estados Unidos | 89 |
| 19 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 3: Policy Gradients | Stanford Online @stanfordonline | — | Estados Unidos | 89 |
| 20 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 13: Meta RL | Stanford Online @stanfordonline | — | Estados Unidos | 89 |
| 5 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 1: Class Intro | Stanford Online @stanfordonline | 84.0K | Estados Unidos | 90 |
| 12 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 4: Actor-Critic Methods | Stanford Online @stanfordonline | 8.3K | Estados Unidos | 89 |
| 15 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 10: RL for LLM Reasoning | Stanford Online @stanfordonline | 5.5K | Estados Unidos | 89 |
| 9 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 17: Advancing Robot Intelligence | Stanford Online @stanfordonline | 5.5K | Estados Unidos | 89 |
| 17 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 8: Reward Learning | Stanford Online @stanfordonline | 4.0K | Estados Unidos | 89 |
| 6 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 11: Model-Based RL | Stanford Online @stanfordonline | 3.2K | Estados Unidos | 90 |
| 8 | Stanford CS221 | Autumn 2025 | Lecture 7: Markov Decision Processes | Stanford Online @stanfordonline | 818 | Estados Unidos | 89 |
| 11 | Stanford CS221 | Autumn 2025 | Lecture 8: Reinforcement Learning | Stanford Online @stanfordonline | 649 | Estados Unidos | 89 |
| 16 | MASTERING MACHINES: The Reinforcement Learning Revolution in AI | TechWaveWeekly @techwaveweekly-zg9jr | 406 | Austria | 89 |
| 4 | Q-Learning | | 259 | — | 90 |
| 14 | An Introduction to Reinforcement Learning | | 247 | Estados Unidos | 89 |
| 1 | Lecture 5.1 | Reinforcement learning | Reinforcement learning in practice | #mlt #aktu #unit5 | Tech Master Edu @techmasteredu | 67 | India | 95 |
| 13 | Reinforcement Learning: How Machines Learn Through Rewards | 머니렙아크 MoneyRepArc @머니렙아크moneyreparc | 0 | Corea del Sur | 89 |