flâneur — a map of the web's best reading

AI Safety | Arkose

victoriabrook.github.io · 1,173 words · saved by 1 readers

AI Safety Resources

AI Safety | Arkose Learn about large-scale risks from advanced AI Selected Papers --> --> --> Overviews What types of risks from advanced AI might we face? These papers provide an overview of anticipated problems and relevant technical research directions. (Ngo et al., 2022) The Alignment Problem from a Deep Learning Perspective (Hendrycks et al., 2023) An Overview of Catastrophic AI Risks (Chan et al., 2023) Harms from Increasingly Agentic Algorithmic Systems (Anwar et al., 2024) Foundational Challenges in Assuring Alignment and Safety of Large Language Models Dangerous Capability Evaluations

Explore this link on the map →

saved by

related reading