P

EPISODE · Apr 14, 2026 · 16 MIN

The Sound of Reasoning: Unpacking NVIDIA's 'Audio Flamingo Next' and the 30-Minute Context Window

from Paper Trail

This episode introduces Audio Flamingo Next (AF-Next), a new AI model from NVIDIA and the University of Maryland, which significantly advances multimodal AI by closing the "audio gap." It explains the inherent difficulties of processing continuous, complex audio compared to discrete text, detailing AF-Next's innovative architecture, including its 30-second chunking strategy and specialized components. Listeners will learn how this generalist model unifies various audio tasks and can understand and reason over extended audio files, outperforming existing systems.

Episode metadata supplied by the publisher feed · Published Apr 14, 2026

Embed this episode

NOW PLAYING

The Sound of Reasoning: Unpacking NVIDIA's 'Audio Flamingo Next' and the 30-Minute Context Window

0:00 16:59

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of Paper Trail?

This episode is 16 minutes long.

When was this Paper Trail episode published?

This episode was published on April 14, 2026.

Can I download this Paper Trail episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!