Problem Areas in Physics and AI Safety | Apart Research
Apart Research is an independent research organization focusing on AI safety. We accelerate AI safety research through mentorship, collaborations, and research sprints
Problem Areas in Physics and AI Safety | Apart Research Problem Areas in Physics and AI Safety We outline five key problem areas in AI safety for the AI Safety x Physics hackathon. Ari Brill Researcher Lauren Greenspan Researcher Title Although there is a lot of good physics for AI literature out there, we take a predominantly ‘problem first’ approach. This is to avoid restricting solutions to specific physics fields, methods, and tools. We’re excited to reframe these old perspectives and find new ones! To guide you in getting started, we’ve developed the following problem space for AI safety,
Explore this link on the map →related reading
- Statistical Physics for Ambitious Interpretability: A Workshop Retrospective — LessWronglesswrong.com
- Dario Amodei — The Urgency of Interpretabilitydarioamodei.com
- Vibe physics: The AI grad student \ Anthropicanthropic.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- Transformer Circuits Threadtransformer-circuits.pub
- AI in 2025: gestalt — LessWronglesswrong.com
- A Pragmatic Vision for Interpretability — AI Alignment Forumalignmentforum.org
- What's the difference -- (physics of) AI, physics, math and interpretability | Ziming Liukindxiaoming.github.io
- ARENA - AI Safety Curriculumlearn.arena.education
- On Optimism for Interpretabilitygoodfire.ai