| ★ | Deep Q Networks | | 64 | — | |
| 4 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 6: Q-Learning | Stanford Online @stanfordonline | — | Estados Unidos | 86 |
| 9 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Tutorial Session: Review of Q-Learning | Stanford Online @stanfordonline | — | Estados Unidos | 85 |
| 19 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 8: Reward Learning | Stanford Online @stanfordonline | 4.0K | Estados Unidos | 84 |
| 16 | Layers of Ai | | 1.7K | — | 84 |
| 8 | Why neural networks are so deep? (AlexNet - Explained) | CodeEmporium @codeemporium | 876 | Estados Unidos | 85 |
| 13 | Reinforcement Learning | Algorithms | ML | Machine Learning | AI | Btech | BSc | Diploma | BCA | | 827 | India | 85 |
| 12 | Neural Networks and Deep Learning (Complete Course) | Nerd's lesson @nerdslesson | 750 | Pakistán | 85 |
| 11 | Temporal Difference Learning | | 670 | — | 85 |
| 5 | What is Reinforcement Learning? | | 581 | — | 86 |
| 15 | Deconvolution - what do networks learn? (visualization + code) | CodeEmporium @codeemporium | 546 | Estados Unidos | 84 |
| 3 | D for Deep Learning | | 524 | India | 87 |
| 14 | Deep Learning with TensorFlow | | 424 | Estados Unidos | 84 |
| 1 | Q-Learning | | 259 | — | 89 |
| 2 | What is Deep Learning (DL)? #artificialintelligence | NobleX Infinity Labs®️ @noblexinfinitylabs | 259 | India | 88 |
| 6 | Lecture 5.2 | Reinforcement learning models | Q Learning | Markov Decision Process | #mlt #aktu #ml | Tech Master Edu @techmasteredu | 250 | India | 86 |
| 20 | An Introduction to Reinforcement Learning | | 247 | Estados Unidos | 84 |
| 18 | Actor-Critic Methods | | 170 | — | 84 |
| 10 | Policy Gradient Methods | | 113 | — | 85 |
| 17 | Nash-DQN: La Inteligencia Artificial que Domina Juegos Complejos | | 11 | — | 84 |
| 7 | Reinforcement Learning: How Machines Learn Through Rewards | 머니렙아크 MoneyRepArc @머니렙아크moneyreparc | 0 | Corea del Sur | 85 |