| ★ | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 8: Reward Learning | Stanford Online @stanfordonline | 4.0K | Estados Unidos | |
| 1 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 5: Off-Policy Actor Critic | Stanford Online @stanfordonline | — | Estados Unidos | 88 |
| 4 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 15: Hierarchical RL and IL | Stanford Online @stanfordonline | — | Estados Unidos | 86 |
| 7 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 6: Q-Learning | Stanford Online @stanfordonline | — | Estados Unidos | 85 |
| 9 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 16: RL for Robots | Stanford Online @stanfordonline | — | Estados Unidos | 82 |
| 10 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 14: Exploration | Stanford Online @stanfordonline | — | Estados Unidos | 82 |
| 13 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 3: Policy Gradients | Stanford Online @stanfordonline | — | Estados Unidos | 81 |
| 19 | Equation for an R Value | eHowEducation @ehoweducation | — | — | 80 |
| 2 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 2: Imitation Learning | Stanford Online @stanfordonline | 15.7K | Estados Unidos | 88 |
| 5 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 4: Actor-Critic Methods | Stanford Online @stanfordonline | 8.3K | Estados Unidos | 86 |
| 16 | Semplificazione tra radicali #matematicaconlidia #matematica #math #radicali | Matematica con Lidia @matematicaconlidia | 6.1K | — | 81 |
| 6 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 7: Offline RL | Stanford Online @stanfordonline | 5.7K | Estados Unidos | 85 |
| 17 | “TAKE IT OFF!” Learn English Phrasal Verbs for Clothes & Shopping | English with Ronnie · EnglishLessons4U with engVid @engvidronnie | 3.9K | Estados Unidos | 81 |
| 8 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 11: Model-Based RL | Stanford Online @stanfordonline | 3.2K | Estados Unidos | 84 |
| 3 | Stanford CS224R Deep Reinforcement Learning | Spring 2025 | Lecture 12: Multi-Task RL | Stanford Online @stanfordonline | 2.4K | Estados Unidos | 87 |
| 14 | Lecture 18: Transmitting Information Reliably over a Noisy Channel & Shannon’s Noisy Coding Theorem | MIT OpenCourseWare @mitocw | 661 | Estados Unidos | 81 |
| 18 | La prise de notes.partie1 | | 530 | Marruecos | 81 |
| 11 | La ECAP ofrece formación gratuita online a los opositores de 15 especialidades | RTVCE @radiotelevisionceuta | 343 | España | 82 |
| 15 | Psychology and ELT - Intermittent Reinforcement | Nick Michelioudakis @mrnickmi | 128 | — | 81 |
| 20 | Propiedades de la multiplicación | Profe Javier Riveros A @profejavierriverosa3204 | 37 | — | 80 |
| 12 | Tuesday 26 - 5 | Mrs. Ruth's Class @mrsruthsclass | 0 | — | 81 |