Dan Hendrycks on Catastrophic AI Risks episode artwork

EPISODE · Nov 3, 2023 · 2H 7M

Dan Hendrycks on Catastrophic AI Risks

from Future of Life Institute Podcast · host Gus Docker

Dan Hendrycks joins the podcast again to discuss X.ai, how AI risk thinking has evolved, malicious use of AI, AI race dynamics between companies and between militaries, making AI organizations safer, and how representation engineering could help us understand AI traits like deception. You can learn more about Dan's work at https://www.safe.ai Timestamps: 00:00 X.ai - Elon Musk's new AI venture 02:41 How AI risk thinking has evolved 12:58 AI bioengeneering 19:16 AI agents 24:55 Preventing autocracy 34:11 AI race - corporations and militaries 48:04 Bulletproofing AI organizations 1:07:51 Open-source models 1:15:35 Dan's textbook on AI safety 1:22:58 Rogue AI 1:28:09 LLMs and value specification 1:33:14 AI goal drift 1:41:10 Power-seeking AI 1:52:07 AI deception 1:57:53 Representation engineering

Episode metadata supplied by the publisher feed · Published Nov 3, 2023

Embed this episode

NOW PLAYING

Dan Hendrycks on Catastrophic AI Risks

0:00 2:07:25

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of Future of Life Institute Podcast?

This episode is 2 hours and 7 minutes long.

When was this Future of Life Institute Podcast episode published?

This episode was published on November 3, 2023.

Can I download this Future of Life Institute Podcast episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!