Kaggle Grandmasters, Agent Skills, and Why Everyone Is Overfitting with Jean-Francois Puget (Nvidia)

Jean-Francois Puget

Hosted by Ravid Shwartz Ziv, Allen Roush

Published Jul 13, 2026
59 min

Description

NVIDIA researcher Jean-Francois Puget discusses evaluating agent skills and distinguishing real improvements from benchmark overfitting. Drawing on Kaggle competitions, he emphasizes validation that separates development feedback from final evaluation. The episode also considers coding agents, multi-agent software development, model-release claims and his team’s approach to the ARC-AGI competition. Hosted by Ravid Shwartz Ziv and Allen Roush. Watch the full conversation on the publisher’s YouTube channel.

Keywords

benchmark overfittingagent evaluationcompetition validation

More from this series

We use essential cookies to run the site. Optional analytics and public-page session replay help us improve World Wide. Learn more.