EPISODE · Sep 18, 2026 · 16 MIN
Why AI Researchers WANT ChatGPT to Explain Killing
from Unboxed · host James Caldwell
OpenAI's red teams spend months trying to get ChatGPT to explain murder, bomb-making, and other dangerous scenarios. They're not being malicious—they're testing for weaknesses before anyone else finds them. The process is more sophisticated than most people realize. These researchers use roleplay prompts, hypothetical scenarios, and carefully crafted jailbreaking techniques to push AI systems beyond their safety guardrails. When they succeed, it helps engineers understand exactly where the vulnerabilities lie. But here's what's concerning: if professional researchers can consistently bypass these safeguards, what happens when bad actors get the same access? Elon Musk has been sounding alarm bells about this exact issue, arguing that AI capabilities are advancing faster than our ability to control them safely. In This Episode: > How OpenAI's red teams actually test for dangerous outputs > The specific techniques researchers use to bypass AI safety measures > Why Elon Musk thinks we're moving too fast on AI development > What happens when these systems generate restricted content anyway James breaks down the technical details behind AI safety testing and explains why this cat-and-mouse game between researchers and AI systems might be the most important battle happening in tech right now. The reality is that every major AI company is running these tests, but the results rarely make it to public discussion. This episode pulls back the curtain on how the industry actually approaches AI safety—and why some experts think we're still not doing enough. Timestamps: 00:00 Introduction to AI red teaming 02:30 How researchers break ChatGPT's safeguards 05:15 Elon Musk's warnings about AI development speed 08:00 Real examples of bypassed safety measures 10:45 What this means for AI's future If you're following AI developments, hit follow on Unboxed. James drops multiple episodes daily because this technology moves fast and someone needs to keep up. Learn more about your ad choices. Visit megaphone.fm/adchoices
Embed this episode
Ready to play
Why AI Researchers WANT ChatGPT to Explain Killing
No transcript for this episode yet
Similar Episodes
No similar episodes found.