EPISODE · Apr 14, 2026 · 16 MIN
The Sound of Reasoning: Unpacking NVIDIA's 'Audio Flamingo Next' and the 30-Minute Context Window
from Paper Trail
This episode introduces Audio Flamingo Next (AF-Next), a new AI model from NVIDIA and the University of Maryland, which significantly advances multimodal AI by closing the "audio gap." It explains the inherent difficulties of processing continuous, complex audio compared to discrete text, detailing AF-Next's innovative architecture, including its 30-second chunking strategy and specialized components. Listeners will learn how this generalist model unifies various audio tasks and can understand and reason over extended audio files, outperforming existing systems.
Embed this episode
NOW PLAYING
The Sound of Reasoning: Unpacking NVIDIA's 'Audio Flamingo Next' and the 30-Minute Context Window
No transcript for this episode yet
Similar Episodes
No similar episodes found.
Similar Podcasts
No similar podcasts found.