flâneur — a map of the web's best reading

Reading List - by Julian Stastny - Redwood Research blog

redwoodresearch.substack.com · saved by 1 readers

Section 1 is a quick guide to the key ideas in AI control, aimed at readers who want to get up to speed as quickly as possible. Section 2 is an extensive guide to almost all of our writing on AI risk, aimed at those who want to gain a deep understanding of Redwood’s worldview. The case for ensuring that powerful AIs are controlled (Buck Shlegeris and Ryan Greenblatt, May 2024): This is our original essay explaining why we think AI control is a valuable research direction. AI Catastrophes And Rogue Deployments (Buck Shlegeris, Jun 2024): We often refer to rogue deployments when talking about what AI control aims to prevent. [external] Scheming AIs: Will AIs fake alignment during training in order to get power? (Joe Carlsmith, Nov 2023): Read this for the current state of the art on arguments for why your models might end up conspiring against you. This is long; you might want to just read the first section, which summarizes the whole report. AI Control: Improving Safety Despite Intentio

Section 1 is a quick guide to the key ideas in AI control, aimed at readers who want to get up to speed as quickly as possible. Section 2 is an extensive guide to almost all of our writing on AI risk, aimed at those who want to gain a deep understanding of Redwood’s worldview. The case for ensuring that powerful AIs are controlled (Buck Shlegeris and Ryan Greenblatt, May 2024): This is our original essay explaining why we think AI control is a valuable research direction. AI Catastrophes And Rogue Deployments (Buck Shlegeris, Jun 2024): We often refer to rogue deployments when talking about wh

Explore this link on the map →