Exposing the Rotten Reality of AI Training Data episode artwork

EPISODE · Dec 31, 2023 · 40 MIN

Exposing the Rotten Reality of AI Training Data

from The Tech Policy Press Podcast

In a report released December 20, 2023, the Stanford Internet Observatory said it had detected more than 1,000 instances of verified child sexual abuse imagery in a significant dataset utilized for training generative AI systems such as Stable Diffusion 1.5. This troubling discovery builds on prior research into the “dubious curation” of large-scale datasets used to train AI systems, and raises concerns that such content may contributed to the capability of AI image generators in producing realistic counterfeit images of child sexual exploitation, in addition to other harmful and biased material. Justin Hendrix spoke the report’s author, Stanford Internet Observatory Chief Technologist David Thiel.

Episode metadata supplied by the publisher feed · Published Dec 31, 2023

Embed this episode

NOW PLAYING

Exposing the Rotten Reality of AI Training Data

0:00 40:16

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of The Tech Policy Press Podcast?

This episode is 40 minutes long.

When was this The Tech Policy Press Podcast episode published?

This episode was published on December 31, 2023.

Can I download this The Tech Policy Press Podcast episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!