Intentionally Designing the Future of AI
Progress in technology typically goes hand-in-hand with progress in fundamental science. Throughout history, understanding of the scientific foundations of our technologies has led to revolutions in the way those technologies are built and deployed. From this perspective, the current revolution in AI is a surprising anomaly: the technology is advancing at a staggering pace, but our understanding of it is not. This gap in understanding is alarming, but means that scientifically we are at an incredible juncture: we have the best chance in history to understand minds - the new minds that we're building in datacenters. These minds live not only in the world we're familiar with - text, images, videos, and so on - but in worlds alien to us like the genome and epigenome, protein folding, quantum chemistry, and materials science. Goodfire's goal is to use interpretability techniques to guide the new minds we're building to share our values, and to learn from them where they have something to t
Intentionally Designing the Future of AI Blog Intentionally designing the future of AI Author Thomas McGrath Published February 5, 2026 Contents The dawn of intentional design What does intentional design enable? What do we need to do to develop intentional design? Developing intentional design responsibly Progress in technology typically goes hand-in-hand with progress in fundamental science. Throughout history, understanding of the scientific foundations of our technologies has led to revolutions in the way those technologies are built and deployed. From this perspective, the current revolut
Explore this link on the map →saved by
related reading
- Dario Amodei — The Urgency of Interpretabilitydarioamodei.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- On Optimism for Interpretabilitygoodfire.ai
- In (highly contingent!) defense of interpretability-in-the-loop ML training — AI Alignment Forumalignmentforum.org
- A Pragmatic Vision for Interpretability — AI Alignment Forumalignmentforum.org
- The Most Forbidden Technique — LessWronglesswrong.com
- Neel Nanda on Mechanistic Interpretability: Progress, Limits, and Paths to Safer AI — EA Forumforum.effectivealtruism.org
- An Ambitious Vision for Interpretability — AI Alignment Forumalignmentforum.org
- How Can Interpretability Researchers Help AGI Go Well? — AI Alignment Forumalignmentforum.org
- The Hitchhiker's Guide to Actionable Interpretabilityactionable-interpretability-guide.github.io
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com