Standards | Guidelight AI Standards
guidelight.ai · 167 words · saved by 1 readers
Guidelight's set of AI standards
Control (v1.1) View full standard → Control refers to the technical and operational measures that constrain what an AI system can do, regardless of whether it is aligned. These measures both reduce catastrophic risk from a misaligned AI and can surface evidence of an AI's misalignment. Alignment (v1.0) View full standard → Alignment is the degree to which an AI system understands how it is intended to behave, tries to behave that way, and reliably does so. This standard covers the methods developers use to improve and measure alignment. Transparency (v1.0) View full standard →…
saved by
related reading
- Standards Development Process | Guidelightguidelight.ai
- A Summary of Recent Work (July 2026)gdmalignment.substack.com
- CAIS AI Dashboarddashboard.safe.ai
- The case for ensuring that powerful AIs are controlled — LessWronglesswrong.com
- Guidelight's Control Assessment of Frontier AI Companiesguidelight.ai
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- Introduction to AI Control - by Sarah - BlueDot Impactblog.bluedot.org
- Thoughts on the conservative assumptions in AI controlblog.redwoodresearch.org
- Seoul Alignment Workshop 2026: What We Learnedfar.ai
- Reading Listblog.redwoodresearch.org
- The case for ensuring that powerful AIs are controlledblog.redwoodresearch.org
- Research Areas in Benchmark Design and Evaluation (The Alignment Project by UK AISI) — AI Alignment Forumalignmentforum.org