Paper Feed
Highlighting research I find interesting and think may deserve more attention (as of 07/06/26). Alignment and Control AI Economics and Forecasting Evaluations and Agents Interpretability Training and Generalization Miscellaneous
Paper Feed: July 2026 Highlighting research I find interesting and think may deserve more attention (as of 08/01/26). Sections: Alignment and Control Diffuse AI Control on Fuzzy Tasks Mikhail Terekhov, Caglar Gulcehre, Vivek Hebbar, Joe Benton (2026) Sub-agent delegation chaining David Rein (2026) Data filtering works a lot worse than you would expect Dohun Lee, J Rosser, Josh Engels, Neel Nanda (2026) Fitness-seeking AIs Alex Mallen (2026) Distributed Attacks in Persistent-State AI Control Josh Hills, Ida Caspary, Asa Cooper Stickland (2026) Distributed Denial of Science: How…
saved by
related reading
- AI 2027ai-2027.com
- LessWronglesswrong.com
- Reading Listblog.redwoodresearch.org
- The Case Against AI Control Research — LessWronglesswrong.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- Introduction to AI Control - by Sarah - BlueDot Impactblog.bluedot.org
- I think alignment work is more promising than control work — LessWronglesswrong.com
- Research Areas in AI Control (The Alignment Project by UK AISI) — AI Alignment Forumalignmentforum.org
- Current AIs seem pretty misaligned to me — LessWronglesswrong.com
- AI 2027ai-2027.com
- AI in 2025: gestalt — LessWronglesswrong.com
- Research Areas in Benchmark Design and Evaluation (The Alignment Project by UK AISI) — AI Alignment Forumalignmentforum.org