✳flâneur — a map of the web's best reading
(My understanding of) What Everyone in Technical Alignment is Doing and Why - LessWrong
lesswrong.com · 20,052 words · saved by 1 readers
Epistemic Effort: ~75 hours of work put into this document …
x (My understanding of) What Everyone in Technical Alignment is Doing and Why — LessWrong Research Agendas AI Alignment Fieldbuilding AI Curated 415 (My understanding of) What Everyone in Technical Alignment is Doing and Why by Thomas Larsen , elifland 29th Aug 2022 AI Alignment Forum 45 min read 90 415 Ω 95 Epistemic Status: My best guess Epistemic Effort: ~75 hours of work put into this document Contributions: Thomas wrote ~85% of this, Eli wrote ~15% and helped edit + structure it. Unless specified otherwise, writing in the first person is by Thomas and so are the opinions. Thanks to Mirand
Explore this link on the map →related reading
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- A Pragmatic Vision for Interpretability — AI Alignment Forumalignmentforum.org
- Shallow review of technical AI safety, 2024 — LessWronglesswrong.com
- A positive case for how we might succeed at prosaic AI alignment — AI Alignment Forumalignmentforum.org
- Automated Alignment Researchers: Using large language models to scale scalable oversight \ Anthropicanthropic.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Externalized reasoning oversight: a research direction for language model alignment — AI Alignment Forumalignmentforum.org
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- How Can Interpretability Researchers Help AGI Go Well? — AI Alignment Forumalignmentforum.org
- Alignment remains a hard, unsolved problem — AI Alignment Forumalignmentforum.org