The Information Bottleneck
INRIA researcher David Holzmüller explains how to choose among tabular foundation models, boosted trees and multilayer perceptrons. He discusses TabArena, TabICL and RealMLP, benchmark weaknesses, small-data performance, calibration and class imbalance. The conversation also examines text columns, explainability, time-series distinctions, synthetic pretraining and the trade-offs among computation, inference speed and predictive performance. Hosted by Ravid Shwartz Ziv and Allen Roush. Watch the full conversation on the publisher’s YouTube channel.