AI incidents, audits, and the limits of benchmarks episode artwork

EPISODE · Feb 13, 2026 · 42 MIN

AI incidents, audits, and the limits of benchmarks

from Practical AI · host Daniel Whitenack and Chris Benson

AI is moving fast from research to real-world deployment, and when things go wrong, the consequences are no longer hypothetical. In this episode, Sean McGregor, co-founder of the AI Verification & Evaluation Research Institute and also the founder of the AI Incident Database, joins Chris and Dan to discuss AI safety, verification, evaluation, and auditing. They explore why benchmarks often fall short, what red-teaming at DEFCON reveals about machine learning risks, and how organizations can better assess and manage AI systems in practice.Featuring:Sean McGregor– LinkedInChris Benson – Website, LinkedIn, Bluesky, GitHub, XDaniel Whitenack – Website, GitHub, XLinks:AI Verification & Evaluation Research InstituteAI Incident Database38th convening of IAAIBenchRiskState of Global AI Incident ReportingUpcoming Events: Register for upcoming webinars here!

Episode metadata supplied by the publisher feed · Published Feb 13, 2026

Embed this episode

Ready to play

AI incidents, audits, and the limits of benchmarks

0:00 42:52

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

Frequently Asked Questions

How long is this episode of Practical AI?

This episode is 42 minutes long.

When was this Practical AI episode published?

This episode was published on February 13, 2026.

Is there a transcript available for this episode?

Yes, a full transcript is available for this episode. You can read the complete transcript on the episode page.

Can I download this Practical AI episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!