Things Read | Sept/Oct 2024 | kipply's blog
A checklist format of what we need to succeed at AI Safety, and it’s specifically safety and not alignment as there’s a pretty good case for a Control agenda. I care a lot about existential threats to humanity or to other things we value, but I also think that the upside from AGI is at least comparable. Mandatory Machines of Loving Grace (rare Dario Amodei public content) which I think is rather tame but alas it seems like everyone is upset that it’s either too tame or too aggressive. 100M context windows, someone please make use of it. sama will never top the American equity fund, but at least the website for the intelligence age is still good. Best tech drama lately? Definitely Automattic. See also this and well, the orange site. SSC vintage on reversing advice which I quite like the underlying model for. Probably I should write some text that is advice on taking advice. We love Baudrillard, and we love Baudrillard gloves! This years Lesswrong Petrov day is iconic as usual. People ar
Things Read | Sept/Oct 2024 Nov 11, 2024 1 Share A checklist format of what we need to succeed at AI Safety , and it’s specifically safety and not alignment as there’s a pretty good case for a Control agenda . I care a lot about existential threats to humanity or to other things we value, but I also think that the upside from AGI is at least comparable. Mandatory Machines of Loving Grace (rare Dario Amodei public content) which I think is rather tame but alas it seems like everyone is upset that it’s either too tame or too aggressive. 100M context windows, someone please make use of it. sama w
Explore this link on the map →related reading
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- The Possessed Machines: Dostoevsky's Demons and the Coming AGI Catastrophepossessedmachines.com
- Where I agree and disagree with Eliezer — LessWronglesswrong.com
- You will be OK — LessWronglesswrong.com
- AGI safety from first principles: Introduction — LessWronglesswrong.com
- AI Safety for Fleshy Humans: a whirlwind touraisafety.dance
- I'm Switching Into AI Safetyalexirpan.com
- Critical review of Christiano's disagreements with Yudkowsky — LessWronglesswrong.com
- AI safety - Wikipediaen.wikipedia.org
- The Best of LessWrong — LessWronglesswrong.com
- Saying Goodbye — LessWronglesswrong.com
- Shtetl-Optimized >> Blog Archive >> Why am I not terrified of AI?scottaaronson.blog