✳flâneur — a map of the web's best reading
Not a Paper: "Frontier Lab CEOs are Capable of In-Context Scheming" — LessWrong
lesswrong.com · 3,616 words · saved by 2 readers
(Fragments from a research paper that will never be written, but whose existence was brought to my attention by GradientDissenter.) …
x Not a Paper: "Frontier Lab CEOs are Capable of In-Context Scheming" — LessWrong Humor AI Frontpage 2026 Top Fifty: 12 % 229 Not a Paper: "Frontier Lab CEOs are Capable of In-Context Scheming" by LawrenceC 29th Apr 2026 8 min read 8 229 (Fragments from a research paper that will never be written, but whose existence was brought to my attention by GradientDissenter .) Extended Abstract. The CEOs of frontier AI developers are becoming increasingly powerful and wealthy, significantly increasing their potential for risks. One concern is that of executive misalignment: when the CEO has different i
Explore this link on the map →saved by
related reading
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- Current AIs seem pretty misaligned to me — LessWronglesswrong.com
- The hot mess theory of AI misalignment: More intelligent agents behave less coherently | Jascha’s blogsohl-dickstein.github.io
- Current AIs seem pretty misaligned to meblog.redwoodresearch.org
- Agentic Misalignment: How LLMs Could be Insider Threats — LessWronglesswrong.com
- Model Organisms of Misalignment: The Case for a New Pillar of Alignment Research — AI Alignment Forumalignmentforum.org
- Agentic misalignment: How LLMs could be insider threats \ Anthropicanthropic.com
- Alignment will happen by default. What’s next? — LessWronglesswrong.com
- Why AI alignment could be hard with modern deep learningcold-takes.com
- Frontier Risk Report (February to March 2026) - METRmetr.org
- Current AIs seem pretty misaligned to me — AI Alignment Forumalignmentforum.org
- 39 - Evan Hubinger on Model Organisms of Misalignment | AXRP - the AI X-risk Research Podcastaxrp.net