RL Was Broken at Every Level - With Joseph Suarez (PufferAI)
Joseph Suarez
Hosted by Ravid Shwartz Ziv, Allen Roush
Published Jul 30, 2026
1 h 3 min
Description
PufferAI researcher Joseph Suarez argues that implementation and simulation bottlenecks have constrained reinforcement learning as much as algorithm choice. He discusses the design of useful simulators, CPU and GPU trade-offs, faster training pipelines and open-source research tools. The conversation ranges from game environments to potential applications in scientific simulation and biological modelling. Hosted by Ravid Shwartz Ziv and Allen Roush. Watch the full conversation on the publisher’s YouTube channel.
Keywords
reinforcement learning systemssimulation throughputscientific simulation
More from this series
- Tiny Recursive Models Beat the Giants - Alexia Jolicoeur-Martineau (Microsoft)Sep 15, 2026 · 52 min
- Why Deep Learning Finally Works on Tables | Frank Hutter (Prior Labs)Aug 24, 2026 · 1 h 18 min
- Surya Ganguli: The Physics of IntelligenceAug 17, 2026 · 1 h 25 min
- Nathan Lambert: Inside Post-Training and the Open Model FightAug 8, 2026 · 1 h 15 min
- Sara Hooker on the End of Static AISep 16, 2026 · 1 h 36 min