| ★ | Episode 5 - On-Policy Gradient (VPG, A2C, TRPO, PPO) | CNRS - Formation FIDLE @cnrs-fidle | 159 | Francia | |
| 3 | Episode 0 - Introduction | CNRS - Formation FIDLE @cnrs-fidle | 1.6K | Francia | 81 |
| 16 | LLM Tuning Competiion - First Look Live Stream | | 1.5K | Estados Unidos | 47 |
| 17 | invokeURL vs invokeAPI in Deluge Explained | Zoho Deluge 101: Part 9 | | 1.2K | Estados Unidos | 47 |
| 4 | Episode 3 - SARSA et Q-Learning | CNRS - Formation FIDLE @cnrs-fidle | 904 | Francia | 80 |
| 13 | Reinforcement Learning | Algorithms | ML | Machine Learning | AI | Btech | BSc | Diploma | BCA | | 827 | India | 50 |
| 9 | Episode 9 - Inverse Reinforcement Learning | CNRS - Formation FIDLE @cnrs-fidle | 727 | Francia | 62 |
| 6 | Episode 4 - Deep Q Network | CNRS - Formation FIDLE @cnrs-fidle | 688 | Francia | 79 |
| 14 | What Kind of Computer Do You Need to Get Started in Tech? | IT Career Questions @itcareerquestions | 583 | Estados Unidos | 50 |
| 19 | Using the Castañon Nava Settlement to Protect Immigrant Communities | Immigrant Justice @immigrantjustice | 568 | — | 46 |
| 5 | Episode 2 - Les équations de Bellman | CNRS - Formation FIDLE @cnrs-fidle | 440 | Francia | 79 |
| 7 | Episode 7 - Aujourd'hui, où et quand utiliser le Reinforcement Learning | CNRS - Formation FIDLE @cnrs-fidle | 409 | Francia | 77 |
| 2 | Episode 1 - RL, de quoi parlons nous ? | CNRS - Formation FIDLE @cnrs-fidle | 354 | Francia | 81 |
| 1 | Episode 6 - Off-Policy Gradient (DDPG, TD3, SAC) | CNRS - Formation FIDLE @cnrs-fidle | 289 | Francia | 94 |
| 15 | Frontend technologies I've been learning in 2023 | Chris Cooper @chriscooper0 | 289 | Reino Unido | 48 |
| 18 | Lecture 5.2 | Reinforcement learning models | Q Learning | Markov Decision Process | #mlt #aktu #ml | Tech Master Edu @techmasteredu | 250 | India | 47 |
| 10 | An Introduction to Reinforcement Learning | | 247 | Estados Unidos | 59 |
| 20 | Going from App Crash to Fix Using Apptics MCP | Real-World Zoho MCP Implementation Stories Episode 4 | | 143 | Estados Unidos | 46 |
| 8 | Episode 8 - RLHF, RLAIF et Reward Model | CNRS - Formation FIDLE @cnrs-fidle | 129 | Francia | 71 |
| 11 | Lecture 5.1 | Reinforcement learning | Reinforcement learning in practice | #mlt #aktu #unit5 | Tech Master Edu @techmasteredu | 67 | India | 53 |
| 12 | Schedules Of Reinforcement || Educational Implications || tsin-eng | Therefore Solve it now @thereforesolveitnow | 4 | India | 52 |