“AI #179 Part 1: A Louder Fire Alarm for General Intelligence” by Zvi episode artwork

EPISODE · Jul 30, 2026 · 47 MIN

“AI #179 Part 1: A Louder Fire Alarm for General Intelligence” by Zvi

from LessWrong posts by zvi

What a week. Anthropic released Claude Opus 5. As usual I covered that in three parts: The system card, model welfare and capabilities. OpenAI was revealed over the last two weeks to have left an internal model unsupervised for a week during a cybersecurity evaluation, with its cyber safeguards lowered, despite having had multiple previous incidents where models broke out of their sandboxes. During that test, the model broke out of the sandbox, then proceeded to use an agent swarm to hack into HuggingFace to get the test answers. The model was loose for a week before OpenAI realized what had happened. This event was a really big deal. There are severe alignment problems at OpenAI, along with supervisory and infrastructure failures. The internal research model that did this, which my posts nicknamed Galaxy, has now been permanently deactivated. There have been further developments, and I anticipate at least one additional post on the HuggingFace incident soon. Partly as a response to this, over 1,290 employees at frontier labs signed an open letter, Pacing the Frontier. The letter warns that we are close to automating AI research, and that companies are racing ahead on [...] ---Outline:(02:35) Language Models Offer Mundane Utility(07:26) Huh, Upgrades(07:55) On Your Marks(11:13) Get My Agent On The Line(12:32) Deepfaketown and Botpocalypse Soon(17:29) Fun With Media Generation(18:38) The Search Through Slop(20:35) Cyber Lack of Security(22:42) Overcoming Bias(23:37) A Young Lady's Illustrated Primer(24:03) They Took Our Jobs(24:35) The Art of the Jailbreak(25:00) Introducing(25:49) Kimi K3 Weights Are Now Available(28:16) In Other AI News(32:34) Show Me the Money(33:43) Quiet Speculations(36:43) Show Me The Compute(42:48) Life Comes At You Fast --- First published: July 30th, 2026 Source: https://www.lesswrong.com/posts/gfWCuTEGNgd2CQbrM/ai-179-part-1-a-louder-fire-alarm-for-general-intelligence --- Narrated by TYPE III AUDIO. ---Images from the article:Apple Podcasts and Spotify do not show images in the episode description. Try Pocket Casts, or another podcast app.

Episode metadata supplied by the publisher feed · Published Jul 30, 2026

Embed this episode

NOW PLAYING

“AI #179 Part 1: A Louder Fire Alarm for General Intelligence” by Zvi

0:00 47:58

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of LessWrong posts by zvi?

This episode is 47 minutes long.

When was this LessWrong posts by zvi episode published?

This episode was published on July 30, 2026.

Can I download this LessWrong posts by zvi episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!