Carlos Guerrero
2 followers · 2 following · 39 views
on the atlas — 14
- On Dwarkesh Patel's Podcast With Ryan Greenblatt4 savers
- An Alien Mind | OpenAI25 savers
- The Origin of Consciousness in the Breakdown of the Bicameral Mind1 savers
- [2206.13353] Is Power-Seeking AI an Existential Risk?1 savers
- The behavioral selection model for predicting AI motivations — LessWrong15 savers
- GPT-6 Astra System Card - OpenAI Deployment Safety Hub1 savers
- Cat-Belling Problems — LessWrong1 savers
- Rationality: A-Z9 savers
- An Argument for Analogies — LessWrong1 savers
- Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident - METR1 savers
- Harsh Bhatt on X: "The Math of RL: From Policy Gradients to PPO and GRPO" / X1 savers
- Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident - METR20 savers
- Policy Gradient Algorithms | Lil'Log8 savers
- [2608.07514] Open Technical Problems in Open-Weight AI Model Risk Management1 savers