Podlipodcast player Webplayer

LessWrong (Curated & Popular)

LessWrong (Curated & Popular)

"“Alignment Engineering” vs. “Misalignment Science”" by Edward James Young

LessWrong (Curated & Popular) · Oct 5, 2026 · 18:06

0:0018:06

Listen in the Podli app 🎧

Follow your favourite podcasts, listen offline and in the car with CarPlay and Android Auto, and always pick up where you left off. Free to try.

There has been much discussion recently around whether a large portion of alignment research is net negative. Without endorsing or refuting them, the basic arguments here are:

On the basis of this argument, some urge alignment researchers at AGI companies to quit outright. But quit to do what? Missing from this exchange so far has been a discussion of opportunity costs. If you aren’t going to do (technical) work on “Alignment” – either inside or outside of an AGI company – what should you work on?

In this post, I outline a contrast between “Alignment Engineering” – the dominant model for what “working on alignment” looks like (inside labs, and in the field as a whole) with “Misalignment Science”. I begin by characterising [...]

---

Outline:

(01:51) "Alignment Engineering"

(06:22) AI Safety and the ML tradition

(07:42) Implicit work trials

(09:36) "Misalignment Science"

(14:56) Conclusion

(15:39) Postscript: Iterating ourselves into oblivion

The original text contained 4 footnotes which were omitted from this narration.

---

First published:
October 5th, 2026

Source:
https://www.lesswrong.com/posts/FogmcDHA6AdMGukum/alignment-engineering-vs-misalignment-science

---



Narrated by TYPE III AUDIO.

Episodes: LessWrong (Curated & Popular)

PodliGet the free Podli app
↓ App