
189: Agentic Loops
Aug 24, 2026 - 82:53
Radio and PodcastLive Radio & Podcasts
Patrick and Jason introduce reinforcement learning and place it alongside supervised and unsupervised learning. They cover Q-learning, SARSA, policy gradients, actor-critic methods, PPO, imitation learning, and why train...