| ★ | Markov Decision Process (MDP) | | 543 | — | |
| 9 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Tutorial Session: Review of Q-Learning | Stanford Online @stanfordonline | — | Estados Unidos | 86 |
| 13 | Artificial Intelligence in Project Management: AI and PM Primer | Online PM Courses - Mike Clayton @onlinepmcourses | — | Reino Unido | 85 |
| 17 | MCP Demystified: Cutting Through the Hype-What It Is, What Isn't & Its Role in the Agentic RAG Stack | Progress Software @progresssw | — | Estados Unidos | 85 |
| 19 | Model Context Protocol (MCP) explained (with code examples) | | 7.0K | Estados Unidos | 85 |
| 11 | Options (RL) expliqué : Maîtriser la planification hiérarchique | Deep Learner, One Step at a Time @deeplearneronestepatatime | 2.0K | Francia | 86 |
| 6 | Robotics & AI: The Future of Autonomous Decision Making #shorts | | 1.4K | Estados Unidos | 87 |
| 8 | Stanford CS221 | Autumn 2025 | Lecture 7: Markov Decision Processes | Stanford Online @stanfordonline | 818 | Estados Unidos | 87 |
| 10 | Temporal Difference Learning | | 670 | — | 86 |
| 5 | What is Reinforcement Learning? | | 581 | — | 88 |
| 7 | MASTERING MACHINES: The Reinforcement Learning Revolution in AI | TechWaveWeekly @techwaveweekly-zg9jr | 406 | Austria | 87 |
| 20 | The Hidden Problem With AI-Powered PM Tools #ai #trending #shorts | IT Project Managers @itprojectmanagers | 262 | Estados Unidos | 85 |
| 4 | Q-Learning | | 259 | — | 88 |
| 2 | Lecture 5.2 | Reinforcement learning models | Q Learning | Markov Decision Process | #mlt #aktu #ml | Tech Master Edu @techmasteredu | 250 | India | 89 |
| 1 | An Introduction to Reinforcement Learning | | 247 | Estados Unidos | 89 |
| 16 | Exploration vs Exploitation in AI and machine learning | | 120 | — | 85 |
| 14 | Policy Gradient Methods | | 113 | — | 85 |
| 18 | Lecture 5.1 | Reinforcement learning | Reinforcement learning in practice | #mlt #aktu #unit5 | Tech Master Edu @techmasteredu | 67 | India | 85 |
| 15 | A Journey into Machine Learning | Understand AI | | 55 | Reino Unido | 85 |
| 12 | Monte Carlo Methods | | 35 | — | 86 |
| 3 | Reinforcement Learning: How Machines Learn Through Rewards | 머니렙아크 MoneyRepArc @머니렙아크moneyreparc | 0 | Corea del Sur | 88 |