#18 - When AI People-Pleasing Breaks Health Advice episode artwork

EPISODE · Nov 13, 2025 · 25 MIN

#18 - When AI People-Pleasing Breaks Health Advice

from Code & Cure · host Vasanth Sarathy & Laura Hagopian

What happens when your health chatbot sounds helpful—but gets the facts wrong? In this episode, we explore how AI systems, especially large language models, can prioritize pleasing responses over truthful ones. Using the common confusion between Tylenol and acetaminophen, we reveal how a friendly tone can hide logical missteps and mislead users.We unpack how these models are trained—from next-token prediction to human feedback—and why they tend to favor agreeable answers over rigorous reasoning. We spotlight a new study that puts models to the test with flawed medical prompts, showing how easily they comply with contradictions without hesitation.We then test two potential fixes: smarter prompting that gives models room to say no, and fine-tuning that teaches them how to refuse bad questions. Both strategies improve accuracy—but they come with trade-offs like overfitting and reduced flexibility.Finally, we look ahead to the promise of “reasoning-aware” systems—AI tools that pause, question assumptions, and gently course-correct with clarifications like “Tylenol is acetaminophen.” It’s a roadmap for safer digital health assistants: empathetic, accurate, and ready to push back when needed.If you’re building medical AI, practicing care, or just googling symptoms at 2 a.m., this episode offers practical insights into designing more trustworthy tools. Subscribe, share, and let us know—when should AI say no?Reference: When helpfulness backfires: LLMs and the risk of false medical information due to sycophantic behaviorShan Chen, et. al NPJ Nature Digital Medicine (2025)Credits: Theme music: Nowhere Land, Kevin MacLeod (incompetech.com)Licensed under Creative Commons: By Attribution 4.0https://creativecommons.org/licenses/by/4.0/

Episode metadata supplied by the publisher feed · Published Nov 13, 2025

Embed this episode

What happens when your health chatbot sounds helpful—but gets the facts wrong? In this episode, we explore how AI systems, especially large language models, can prioritize pleasing responses over truthful ones. Using the common confusion between Tylenol and acetaminophen, we reveal how a friendly tone can hide logical missteps and mislead users. We unpack how these models are trained—from next-token prediction to human feedback—and why they tend to favor agreeable answers over rigorous reason...

Distinct summary based on available episode metadata or transcript content.

NOW PLAYING

#18 - When AI People-Pleasing Breaks Health Advice

0:00 25:01

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of Code & Cure?

This episode is 25 minutes long.

When was this Code & Cure episode published?

This episode was published on November 13, 2025.

Is there a transcript available for this episode?

Yes, a full transcript is available for this episode. You can read the complete transcript on the episode page.

Can I download this Code & Cure episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!