Why TTS Models Now Look Like LLMs — Samuel Humeau, Mistral episode artwork

EPISODE · May 9, 2026

Why TTS Models Now Look Like LLMs — Samuel Humeau, Mistral

from Le Peertube de Tonton · host AI

The dominant architecture pattern for text-to-speech in 2026 looks a lot like an LLM — an autoregressive transformer generating sequences of tokens, one frame of audio at a time. Samuel Humeau from Mistral walks through why the field converged the...

Episode metadata supplied by the publisher feed · Published May 9, 2026

Embed this episode

Ready to play

Why TTS Models Now Look Like LLMs — Samuel Humeau, Mistral

0:00 0:00

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar podcasts found.

Frequently Asked Questions

When was this Le Peertube de Tonton episode published?

This episode was published on May 9, 2026.

Can I download this Le Peertube de Tonton episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!