Why Deep Learning Finally Works on Tables | Frank Hutter (Prior Labs) episode artwork

EPISODE · Aug 24, 2026 · 1H 18M

Why Deep Learning Finally Works on Tables | Frank Hutter (Prior Labs)

from The Information Bottleneck · host Ravid Shwartz-Ziv & Allen Roush

In this episode, Frank Hutter joins us to talk about TabPFN and why tabular data is suddenly the hottest problem in deep learning. Frank is a professor at the University of Freiburg and spent 15 years building the AutoML field before founding Prior Labs, which SAP just acquired for over a billion dollars.We get into why deep learning failed on tables for a decade and what in-context learning changed, how TabPFN is trained entirely on synthetic data, and why a model that never saw a real time series ended up beating specialized forecasting models. Frank also explains the architecture tricks behind scaling from 10,000 to a million rows, where LLMs fit into data science (and where they embarrassingly don't), and what happens to XGBoost from here.Beyond the research, Frank talks about the jump from professor to co-CEO, why he refused to merge his 45-person team into SAP's 110,000 employees, the open-weights licensing debate, and the case for building a frontier lab in Freiburg rather than San Francisco.key topicsThe role of foundation models in tabular dataImpact of SAP acquisition on Pro LabsThe evolution of AutoML and hyperparameter optimizationChallenges and solutions for large context in modelsOpen source models and licensing strategiesThe importance of independence for startup agilityFuture directions in AI for science and medicine00:00 Intro00:34 The SAP acquisition and staying independent07:39 Why tabular data is the next big thing in deep learning14:19 What makes tabular data hard19:14 AutoML, AutoGluon, and fifteen years of hyperparameter tuning28:27 Scaling TabPFN: context limits and architectures34:35 Agentic data science and LLMs39:30 Online learning, time series, and Bayesian inference in a forward pass47:05 Open weights and the license debate54:51 Will LLMs and tabular models merge?1:00:01 From academia to startup1:09:42 Why build in Europe1:12:53 Audience questions and hiringMusic"Kid Kodi" - Blue Dot Sessions - via Free Music Archive - CC BY-NC 4.0.

Episode metadata supplied by the publisher feed · Published Aug 24, 2026

Embed this episode

NOW PLAYING

Why Deep Learning Finally Works on Tables | Frank Hutter (Prior Labs)

0:00 1:18:06

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of The Information Bottleneck?

This episode is 1 hour and 18 minutes long.

When was this The Information Bottleneck episode published?

This episode was published on August 24, 2026.

Can I download this The Information Bottleneck episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!