✳flâneur — a map of the web's best reading
The Parable of the Prinia’s Egg: An Allegory for AI Science | Naomi Saphra
nsaphra.net · 2,040 words · saved by 3 readers
I discuss what counts as strong evidence for an explanation of model behavior.
The Parable of the Prinia's Egg: An Allegory for AI Science Sep 17, 2023 · Naomi Saphra · 10 min read When European scientists first encountered the eggs of the tawny-flanked prinia Prinia subflava , an African nesting bird, they believed they understood what they had found. The prinia lay eggs that exhibited swirls, speckles, and coloration unique to each individual bird. What purpose do such patterns serve in nature? Surely, scientists agreed, these markings allowed the eggs to camouflage and blend into the nest. One 19th century naturalist, Charles Francis Massey Swynnerton, offered an alte
Explore this link on the map →saved by
related reading
- Dario Amodei — The Urgency of Interpretabilitydarioamodei.com
- Interpretability Creationism | Naomi Saphransaphra.github.io
- A Pragmatic Vision for Interpretability — AI Alignment Forumalignmentforum.org
- An Extremely Opinionated Annotated List of My Favourite Mechanistic Interpretability Papers v2 — AI Alignment Forumalignmentforum.org
- How Can Interpretability Researchers Help AGI Go Well? — AI Alignment Forumalignmentforum.org
- On Optimism for Interpretabilitygoodfire.ai
- Your Model Organisms Might Be Fried — LessWronglesswrong.com
- Model Organisms of Misalignment: The Case for a New Pillar of Alignment Research — AI Alignment Forumalignmentforum.org
- Neel Nanda on Mechanistic Interpretability: Progress, Limits, and Paths to Safer AI — EA Forumforum.effectivealtruism.org
- An Extremely Opinionated Annotated List of My Favourite Mechanistic Interpretability Papers v2 — AI Alignment Forumalignmentforum.org
- Why I'm Moving from Mechanistic to Prosaic Interpretability — LessWronglesswrong.com
- Perfectly Normalperfectlynormal.co.uk