EPISODE · Jul 23, 2026 · 23 MIN
Testing the Unpredictable: QA for Agentic AI
from My Weird Prompts
Traditional software testing assumes determinism — run the same test, get the same result. Agentic AI shatters that assumption. This episode maps the emerging QA landscape for probabilistic systems: from golden datasets and LLM-as-judge to trajectory evaluation and adversarial prompting suites like Garak. We explore what carries over from traditional testing, what requires entirely new methodologies, and what a sane minimum testing stack looks like for teams shipping agentic systems. Episode #238681 — open it directly at myweirdprompts.com/238681
Embed this episode
NOW PLAYING
Testing the Unpredictable: QA for Agentic AI
No transcript for this episode yet
Similar Episodes
No similar episodes found.
Similar Podcasts
No similar podcasts found.