“Notes on technical alignment via human-like social drives” by Steven Byrnes episode artwork

EPISODE · Jul 8, 2026 · 1H 22M

“Notes on technical alignment via human-like social drives” by Steven Byrnes

from LessWrong (30+ Karma)

1. Frontmatter 1.1 Backstory for this post As my regular readers know (see Intro to Brain-Like-AGI Safety), I’m working on the technical alignment problem for a hypothetical future “brain-like AGI”, with a particular focus on how human social and moral drives work. After all, if it's possible for humans to do stuff that ultimately leads to a good future, then it's probably also possible for sufficiently human-like AGIs to do stuff that ultimately leads to a good future. Or if it's not possible for humans to do stuff that ultimately leads to a good future, then we’re screwed no matter what. But assuming it's possible, the “sufficiently human-like AGIs” would certainly need to have good prosocial motivations. This is an unsolved problem, and very much not the default (see We need a field of Reward Function Design), but there's probably some solution that's inspired by how humans (sometimes) wind up with good prosocial motivations. I’ve been working on this problem for years, but most of that work has involved laying foundations (e.g. trying to understand how human social drives work). Whereas in the past four months, I’ve been thinking very directly about how to apply those ideas to AGI. [...] The original text contained 13 footnotes which were omitted from this narration. --- First published: July 8th, 2026 Source: https://www.lesswrong.com/posts/rKdS7i4StaMmFzYRo/notes-on-technical-alignment-via-human-like-social-drives --- Narrated by TYPE III AUDIO. ---Images from the article:Apple Podcasts and Spotify do not show images in the episode description. Try Pocket Casts, or another podcast app.

Episode metadata supplied by the publisher feed · Published Jul 8, 2026

Embed this episode

NOW PLAYING

“Notes on technical alignment via human-like social drives” by Steven Byrnes

0:00 1:22:06

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar episodes found.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of LessWrong (30+ Karma)?

This episode is 1 hour and 22 minutes long.

When was this LessWrong (30+ Karma) episode published?

This episode was published on July 8, 2026.

Can I download this LessWrong (30+ Karma) episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!