flâneur — a map of the web's best reading

SLT for AI Safety – Jesse Hoogland

jessehoogland.com · 1,023 words · saved by 1 readers

Jesse Hoogland's personal website

SLT for AI Safety – Jesse Hoogland SLT for AI Safety This sequence draws from a position paper co-written with Simon Pepin Lehalleur, Jesse Hoogland, Matthew Farrugia-Roberts, Susan Wei, Alexander Gietelink Oldenziel, Stan van Wingerden, George Wang, Zach Furman, Liam Carroll, Daniel Murfet. Thank you to Stan, Dan, and Simon for providing feedback on this post. **Alignment **\(\subseteq\) Capabilities. As of 2025, there is essentially no difference between the methods we use to align models and the methods we use to make models more capable. Everything is based on deep learning, and the main d

Explore this link on the map →

saved by

related reading