AI Companies Are Hiring Philosophers and Uncensored AI Faces a Legal Test episode artwork

EPISODE · Jul 6, 2026 · 30 MIN

AI Companies Are Hiring Philosophers and Uncensored AI Faces a Legal Test

from They Might Be Self-Aware · host Hunter Powers, Daniel Bishop, Gary

AI companies are hiring philosophers, and researchers just tested uncensored AI on Swiss Supreme Court cases the big chatbots refuse to read. The philosophy degree is in perhaps the highest demand it has ever been, and the companies doing the hiring are OpenAI, Anthropic, and the rest of the AI labs. Philosophy majors are now more likely to be employed after college than computer science graduates. So what does a philosopher actually do at an AI company? Hunter Powers and Daniel Bishop land on AI alignment as the real job description, and on why Hunter keeps calling that alignment brainwashing. Along the way: soul.md files (the personality document some AI assistants now ship with), whether a soul is just a system prompt with better marketing, Anthropic's 78-page AI constitution, models that notice when they are being tested, and the case for treating AI like a very smart high schooler. Plus a note on Anthropic's Claude Fable model, which some users report they can finally access again. The second half is TF-RefusalBench, a new multilingual benchmark built from real Swiss Federal Supreme Court criminal rulings in German, French, Italian, and English. The finding: aligned models refuse legitimate legal work. OpenAI's open weight model GPT-OSS refused the most, Google's Gemma answered but wrapped everything in warnings, and Qwen mostly just did the work. Which drags the show into uncensored AI. Abliterated models (sometimes called obliterated) have the refusal edited directly out of their weights, and the episode gets into who actually wants them: the role players every model description politely nods to, the investors chasing advice Claude and ChatGPT will not give, and the defense lawyer who genuinely needs a model to read 2,000 pages about a murder. The catch the researchers found: force a model to never say no and it also becomes less truthful. How much of a soul is just the ability to say no? They Might Be Self-Aware is the AI podcast from The Blur, reported from inside the dissolving line between human and machine, not from a safe distance. CHAPTERS 0:00 Cold Open (Gary's Intro) 1:55 Soul.md vs System Prompts 4:25 AI Companies Hiring Philosophers 8:12 AI Alignment as Brainwashing 13:34 AI, the Smart High Schooler 16:42 TF-RefusalBench (Swiss Courts) 20:07 Uncensored AI Models 25:45 AI's Right to Say No 28:28 Uncensored AI Is Less Truthful 29:34 Philosophy Majors Beat CS Grads LISTEN / WATCH EVERYWHERE 🎧 Apple Podcasts: https://podcasts.apple.com/us/podcast/they-might-be-self-aware/id1730993297 🎧 Spotify: https://open.spotify.com/show/3EcvzkWDRFwnmIXoh7S4Mb?si=3d0f8920382649cc 🎧 Everywhere else plus episode page: https://theblur.ai THE BLUR Follow: @TheBlurAI COMMENT Give us the one question an AI should always refuse to answer, no matter who is asking. Or convince us there is no such question. You're listening to They Might Be Self-Aware, from The Blur. New episodes Monday and Thursday. #AI #TMBSA #UncensoredAI #AIAlignment #PhilosophyMajors

Episode metadata supplied by the publisher feed · Published Jul 6, 2026

Embed this episode

AI companies are hiring philosophers, and researchers just tested uncensored AI on Swiss Supreme Court cases the big chatbots refuse to read. Hunter Powers and Daniel Bishop take the philosopher hiring story seriously: why OpenAI and Anthropic suddenly want philosophy majors (now more likely to be employed than computer science grads), and why AI alignment keeps looking like brainwashing a very smart high schooler that knows it is being tested. Then TF-RefusalBench, a benchmark built from real Swiss Federal Supreme Court criminal rulings, catches aligned models refusing legitimate legal work, and the workaround is uncensored, abliterated AI models with the refusals edited out of their weights. The researchers' catch: a model that cannot say no also gets less truthful. Torture, it turns out, does not work on people or on large language models.

Distinct summary based on available episode metadata or transcript content.

NOW PLAYING

AI Companies Are Hiring Philosophers and Uncensored AI Faces a Legal Test

0:00 30:40

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of They Might Be Self-Aware?

This episode is 30 minutes long.

When was this They Might Be Self-Aware episode published?

This episode was published on July 6, 2026.

Can I download this They Might Be Self-Aware episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!