| ★ | Actor-Critic Methods | | 170 | — | |
| 11 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 4: Actor-Critic Methods | Stanford Online @stanfordonline | 8.3K | Estados Unidos | 86 |
| 15 | The 5Ps Framework for Agentic Development | | 4.3K | Estados Unidos | 85 |
| 12 | Options (RL) expliqué : Maîtriser la planification hiérarchique | Deep Learner, One Step at a Time @deeplearneronestepatatime | 2.0K | Francia | 85 |
| 13 | Reinforcement Learning: Zero to Hero | CodeEmporium @codeemporium | 1.4K | Estados Unidos | 85 |
| 2 | Temporal Difference Learning | | 670 | — | 87 |
| 5 | What is Reinforcement Learning? | | 581 | — | 87 |
| 18 | Markov Decision Process (MDP) | | 543 | — | 85 |
| 9 | MASTERING MACHINES: The Reinforcement Learning Revolution in AI | TechWaveWeekly @techwaveweekly-zg9jr | 406 | Austria | 86 |
| 14 | Q-learning with Flow-Matching Policies | Microsoft Research @microsoftresearch | 363 | Estados Unidos | 85 |
| 19 | Lecture 4.1 | Artificial neural network | Perceptron, Gradient Descent, Delta rule | #mlt #aktu #ml | Tech Master Edu @techmasteredu | 319 | India | 85 |
| 3 | Q-Learning | | 259 | — | 87 |
| 7 | Lecture 5.2 | Reinforcement learning models | Q Learning | Markov Decision Process | #mlt #aktu #ml | Tech Master Edu @techmasteredu | 250 | India | 86 |
| 8 | An Introduction to Reinforcement Learning | | 247 | Estados Unidos | 86 |
| 1 | Policy Gradient Methods | | 113 | — | 91 |
| 4 | Machine Learning Techniques for Agents (10 Minutes) | Technology Whisper @technologywhisper | 102 | Singapur | 87 |
| 10 | Lecture 5.1 | Reinforcement learning | Reinforcement learning in practice | #mlt #aktu #unit5 | Tech Master Edu @techmasteredu | 67 | India | 86 |
| 20 | Cross-Entropy Loss Explained: The Complete Guide for Machine Learning | THE FACT FACTORY @thefactfactoryf | 37 | — | 85 |
| 17 | Elevate Your AI: Key Strategies for Continuous Improvement and Ethical Development | | 34 | Estados Unidos | 85 |
| 16 | Why Weight Decay IS L2 Regularization: Complete Guide Explained | THE FACT FACTORY @thefactfactoryf | 5 | — | 85 |
| 6 | Reinforcement Learning: How Machines Learn Through Rewards | 머니렙아크 MoneyRepArc @머니렙아크moneyreparc | 0 | Corea del Sur | 87 |