“Problems I’ve Tried To Legibilize” By Wei Dai LessWrong (Curated & Popular) podcast

“Problems I’ve Tried to Legibilize” by Wei Dai

6d ago 4:17

Content provided by LessWrong. All podcast content including episodes, graphics, and podcast descriptions are uploaded and provided directly by LessWrong or their podcast platform partner. If you believe someone is using your copyrighted work without your permission, you can follow the process outlined here https://podcastplayer.com/legal.

Looking back, it appears that much of my intellectual output could be described as legibilizing work, or trying to make certain problems in AI risk more legible to myself and others. I've organized the relevant posts and comments into the following list, which can also serve as a partial guide to problems that may need to be further legibilized, especially beyond LW/rationalists, to AI researchers, funders, company leaders, government policymakers, their advisors (including future AI advisors), and the general public.

Philosophical problems
1. Probability theory
2. Decision theory
3. Beyond astronomical waste (possibility of influencing vastly larger universes beyond our own)
4. Interaction between bargaining and logical uncertainty
5. Metaethics
6. Metaphilosophy: 1, 2
Problems with specific philosophical and alignment ideas
1. Utilitarianism: 1, 2
2. Solomonoff induction
3. "Provable" safety
4. CEV
5. Corrigibility
6. IDA (and many scattered comments)
7. UDASSA
8. UDT
Human-AI safety (x- and s-risks arising from the interaction between human nature and AI design)
1. Value differences/conflicts between humans
2. “Morality is scary” (human morality is often the result of status games amplifying random aspects of human value, with frightening results)
3. [...]

---
First published:
November 9th, 2025
Source:
https://www.lesswrong.com/posts/7XGdkATAvCTvn4FGu/problems-i-ve-tried-to-legibilize
---
Narrated by TYPE III AUDIO.

683 episodes

Philosophical problems
1. Probability theory
2. Decision theory
3. Beyond astronomical waste (possibility of influencing vastly larger universes beyond our own)
4. Interaction between bargaining and logical uncertainty
5. Metaethics
6. Metaphilosophy: 1, 2
Problems with specific philosophical and alignment ideas
1. Utilitarianism: 1, 2
2. Solomonoff induction
3. "Provable" safety
4. CEV
5. Corrigibility
6. IDA (and many scattered comments)
7. UDASSA
8. UDT
Human-AI safety (x- and s-risks arising from the interaction between human nature and AI design)
1. Value differences/conflicts between humans
2. “Morality is scary” (human morality is often the result of status games amplifying random aspects of human value, with frightening results)
3. [...]

---
First published:
November 9th, 2025
Source:
https://www.lesswrong.com/posts/7XGdkATAvCTvn4FGu/problems-i-ve-tried-to-legibilize
---
Narrated by TYPE III AUDIO.

Podcasts Worth a Listen

LessWrong (Curated & Popular) « »
“Problems I’ve Tried to Legibilize” by Wei Dai