Skye
0 followers · 265 views
on the atlas — 8
- Want to upskill in technical AI safety? Here are 67 useful resources | 80,000 Hours3 savers
- Deep Reinforcement Learning from Human Preferences1 savers
- Simulators - LessWrong16 savers
- A Big Little Idea Called Legibility15 savers
- The behavioral selection model for predicting AI motivations — LessWrong15 savers
- Tips for Empirical Alignment Research — AI Alignment Forum14 savers
- Eliezer's Unteachable Methods of Sanity — LessWrong11 savers
- Concrete research ideas on AI personas — LessWrong2 savers