Why Speech-to-Text Still Fails at Its Own Name episode artwork

EPISODE · Jul 15, 2026 · 21 MIN

Why Speech-to-Text Still Fails at Its Own Name

from My Weird Prompts

When OpenAI's Whisper transcribed its own name as "Wispr," it exposed the fundamental flaw in speech-to-text: models that hear perfectly but understand nothing. This episode unpacks why homophone errors, dropped negations, and hallucinated punctuation survive even low word-error rates — and explores two competing architectural solutions. We compare the two-pass pipeline (Whisper + LLM cleanup) against unified multimodal models like GPT-4o that process audio and reasoning in a single pass. Which approach actually eliminates the need for human review? And what are the hidden failure modes of each? If you dictate more than a few hundred words a day, this episode will change how you think about voice input. Episode #494803 — open it directly at myweirdprompts.com/494803

Episode metadata supplied by the publisher feed · Published Jul 15, 2026

Embed this episode

NOW PLAYING

Why Speech-to-Text Still Fails at Its Own Name

0:00 21:40

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of My Weird Prompts?

This episode is 21 minutes long.

When was this My Weird Prompts episode published?

This episode was published on July 15, 2026.

Can I download this My Weird Prompts episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!