“Notes on technical alignment via human-like social drives” by Steven Byrnes
“Notes on technical alignment via human-like social drives” by Steven Byrnes  
Podcast: LessWrong (30+ Karma)
Published On: Wed Jul 08 2026
Description: 1. Frontmatter 1.1 Backstory for this post As my regular readers know (see Intro to Brain-Like-AGI Safety), I’m working on the technical alignment problem for a hypothetical future “brain-like AGI”, with a particular focus on how human social and moral drives work. After all, if it's possible for humans to do stuff that ultimately leads to a good future, then it's probably also possible for sufficiently human-like AGIs to do stuff that ultimately leads to a good future. Or if it's not possible for humans to do stuff that ultimately leads to a good future, then we’re screwed no matter what. But assuming it's possible, the “sufficiently human-like AGIs” would certainly need to have good prosocial motivations. This is an unsolved problem, and very much not the default (see We need a field of Reward Function Design), but there's probably some solution that's inspired by how humans (sometimes) wind up with good prosocial motivations. I’ve been working on this problem for years, but most of that work has involved laying foundations (e.g. trying to understand how human social drives work). Whereas in the past four months, I’ve been thinking very directly about how to apply those ideas to AGI. [...] The original text contained 13 footnotes which were omitted from this narration. --- First published: July 8th, 2026 Source: https://www.lesswrong.com/posts/rKdS7i4StaMmFzYRo/notes-on-technical-alignment-via-human-like-social-drives --- Narrated by TYPE III AUDIO. ---Images from the article:Apple Podcasts and Spotify do not show images in the episode description. Try Pocket Casts, or another podcast app.