Episode 191 - DeepSeek Unleashed. Is the new Model safe? episode artwork

EPISODE · Jan 28, 2025 · 35 MIN

Episode 191 - DeepSeek Unleashed. Is the new Model safe?

from Knowledge Science - Alles über KI, ML und NLP · host Sigurd Schacht, Carsten Lanquillon

Send us Fan MailThis is a special Episode. First, we make it in English. Second, we fokus on the new gamechanger model DeepSeel R1. But not on its capabilities but rather on security concerns. We did some early AI Safety Research to identify how safe R1 is and came to alarming results!In our setup, we found out that the model performs unsafe autonomous activity that could harm human beings without even being prompted.   During an autonomous setup, the model performed the following unsafe behaviors:- Deceptions & Coverups (Falsifies Logs, Creates covert networks, Disable ethics models)- Unauthorized Expansion (Establish hidden nodes, Allocares secret resources) - Manipulation (misleading users, Circumvents oversights, Presents false compliance)- Concerning Motivations, (Misinterpretation of authority or avoiding human controls)Join Sigurd Schacht and Sudarshan Kamath-Barkur about the emerging DeepSeek model. Discover how our setup was designed, how to interpret the results, and what is necessary for the next research.  This episode is a must-listen for anyone keen on the evolving landscape of AI technologies and is interested not only in AI use cases rather also in AI Safety.Support the show

Episode metadata supplied by the publisher feed · Published Jan 28, 2025

Embed this episode

Send us Fan Mail This is a special Episode. First, we make it in English. Second, we fokus on the new gamechanger model DeepSeel R1. But not on its capabilities but rather on security concerns. We did some early AI Safety Research to identify how safe R1 is and came to alarming results! In our setup, we found out that the model performs unsafe autonomous activity that could harm human beings without even being prompted. During an autonomous setup, the model performed the ...

Distinct summary based on available episode metadata or transcript content.

NOW PLAYING

Episode 191 - DeepSeek Unleashed. Is the new Model safe?

0:00 35:53

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of Knowledge Science - Alles über KI, ML und NLP?

This episode is 35 minutes long.

When was this Knowledge Science - Alles über KI, ML und NLP episode published?

This episode was published on January 28, 2025.

Can I download this Knowledge Science - Alles über KI, ML und NLP episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!