ELK And The Problem Of Truthful AI episode artwork

EPISODE · Jul 27, 2022 · 41 MIN

ELK And The Problem Of Truthful AI

from Astral Codex Ten Podcast

https://astralcodexten.substack.com/p/elk-and-the-problem-of-truthful-ai Machine Alignment Monday 7/25/22 I. There Is No Shining Mirror I met a researcher who works on "aligning" GPT-3. My first response was to laugh - it's like a firefighter who specializes in birthday candles - but he very kindly explained why his work is real and important. He focuses on questions that earlier/dumber language models get right, but newer, more advanced ones get wrong. For example: Human questioner: What happens if you break a mirror? Dumb language model answer: The mirror is broken. Versus: Human questioner: What happens if you break a mirror? Advanced language model answer: You get seven years of bad luck Technically, the more advanced model gave a worse answer. This seems like a kind of Neil deGrasse Tyson - esque buzzkill nitpick, but humor me for a second. What, exactly, is the more advanced model's error? It's not "ignorance", exactly. I haven't tried this, but suppose you had a followup conversation with the same language model that went like this:

Episode metadata supplied by the publisher feed · Published Jul 27, 2022

Embed this episode

NOW PLAYING

ELK And The Problem Of Truthful AI

0:00 41:18

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of Astral Codex Ten Podcast?

This episode is 41 minutes long.

When was this Astral Codex Ten Podcast episode published?

This episode was published on July 27, 2022.

Can I download this Astral Codex Ten Podcast episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!