flâneur — a map of the web's best reading

My AI Model Delta Compared To Yudkowsky — LessWrong

lesswrong.com · 15,339 words · saved by 1 readers

I don’t natively think in terms of cruxes. But there’s a similar concept which is more natural for me, which I’ll call a delta. Imagine that you and I each model the world (or some part of it) as implementing some program. Very oversimplified example: if I learn that e.g. it’s cloudy today, that means the “weather” variable in my program at a particular time[1] takes on the value “cloudy”. Now, suppose your program and my program are exactly the same, except that somewhere in there I think a certain parameter has value 5 and you think it has value 0.3. Even though our programs differ in only that one little spot, we might still expect very different values of lots of variables during execution - in other words, we might have very different beliefs about lots of stuff in the world. If your model and my model differ in that way, and we’re trying to discuss our different beliefs, then the obvious useful thing-to-do is figure out where that one-parameter difference is. That’s a delta: one

x My AI Model Delta Compared To Yudkowsky — LessWrong AI Risk Natural Abstraction AI Curated 279 My AI Model Delta Compared To Yudkowsky by johnswentworth 10th Jun 2024 5 min read 107 279 Preamble: Delta vs Crux I don’t natively think in terms of cruxes . But there’s a similar concept which is more natural for me, which I’ll call a delta. Imagine that you and I each model the world (or some part of it) as implementing some program. Very oversimplified example: if I learn that e.g. it’s cloudy today, that means the “weather” variable in my program at a particular time [1] takes on the value “cl

Explore this link on the map →

related reading