This Deep Research Agent Ignored the Benchmark and Still Won episode artwork

EPISODE · Jan 1, 2026 · 29 MIN

This Deep Research Agent Ignored the Benchmark and Still Won

from YAAP (Yet Another AI Podcast) · host AI21

Tavily built a Deep Research Agent with production in mind. Something they could actually scale. So they did the unsexy work. They went through millions of agent logs, found where tokens were being wasted, and optimized each section of the system. The result surprised them: they cut token consumption by more than half (!), then tested quality and discovered they topped the DeepResearch Bench without even trying. In this YAAP episode, Yuval sits down with Dean from Tavily to break down how they built it, what they did differently from the usual top approaches, and which design choices made better results possible with far fewer tokens. What you’ll learn: How to reduce token burn without tanking quality Why reading millions of logs beats chasing the flashiest tech The design choices that pushed quality up while tokens dropped hard

Episode metadata supplied by the publisher feed · Published Jan 1, 2026

Embed this episode

Ready to play

This Deep Research Agent Ignored the Benchmark and Still Won

0:00 29:59

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of YAAP (Yet Another AI Podcast)?

This episode is 29 minutes long.

When was this YAAP (Yet Another AI Podcast) episode published?

This episode was published on January 1, 2026.

Can I download this YAAP (Yet Another AI Podcast) episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!