My AI Model Delta Compared To Yudkowsky — LessWrong
I don’t natively think in terms of cruxes. But there’s a similar concept which is more natural for me, which I’ll call a delta. Imagine that you and I each model the world (or some part of it) as implementing some program. Very oversimplified example: if I learn that e.g. it’s cloudy today, that means the “weather” variable in my program at a particular time[1] takes on the value “cloudy”. Now, suppose your program and my program are exactly the same, except that somewhere in there I think a certain parameter has value 5 and you think it has value 0.3. Even though our programs differ in only that one little spot, we might still expect very different values of lots of variables during execution - in other words, we might have very different beliefs about lots of stuff in the world. If your model and my model differ in that way, and we’re trying to discuss our different beliefs, then the obvious useful thing-to-do is figure out where that one-parameter difference is. That’s a delta: one
x My AI Model Delta Compared To Yudkowsky — LessWrong AI Risk Natural Abstraction AI Curated 279 My AI Model Delta Compared To Yudkowsky by johnswentworth 10th Jun 2024 5 min read 107 279 Preamble: Delta vs Crux I don’t natively think in terms of cruxes . But there’s a similar concept which is more natural for me, which I’ll call a delta. Imagine that you and I each model the world (or some part of it) as implementing some program. Very oversimplified example: if I learn that e.g. it’s cloudy today, that means the “weather” variable in my program at a particular time [1] takes on the value “cl
Explore this link on the map →related reading
- AI 2027ai-2027.com
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- trees are harlequins, words are harlequins - the voidnostalgebraist.tumblr.com
- Andrej Karpathy — AGI is still a decade awaydwarkesh.com
- AI 2027ai-2027.com
- World-Model Interpretability Is All We Need — AI Alignment Forumalignmentforum.org
- A positive case for how we might succeed at prosaic AI alignment — AI Alignment Forumalignmentforum.org
- AI #24: Week of the Podcast — LessWronglesswrong.com
- Unfalsifiable stories of doom | Mechanize, Inc.mechanize.work
- 2023 - by Dean W. Ball - Hyperdimensionalhyperdimensional.co
- The Best of LessWrong — LessWronglesswrong.com
- 1a3orn's Shortform — LessWronglesswrong.com