Andi Peng
I am a Co-Founder of humans&. I was previously a researcher at Anthropic working on RL and post-training of Claude 3.5 through 4.5 and before that, a PhD student at MIT CSAIL advised by Jacob Andreas and Julie Shah. I’ve spent summers at Facebook AI Research (FAIR), Boston Dynamics AI Institute, MIT-IBM Watson AI Lab, and before grad school, two years as an AI Resident at Microsoft Research. I did my undergrad at Yale, where I got my start in research with Brian Scassellati and read a lot of dead philosophers. I’m interested in building agents that learn representations from rich human knowledge, whether directly (e.g. from users) or through priors (e.g. from LMs). Currently, I’m thinking a lot about how to utilize pretrained models in conjunction with human feedback to interactively learn aligned preferences/rewards. A history buff at heart, I care deeply about working with non-academic communities to create safe, ethical, and equitable AI. I currently serve as a Special Government Em
Andi Peng Search I am a Co-Founder of humans& . I was previously a researcher at Anthropic working on RL and post-training of Claude 3.5 through 4.5 and before that, a PhD student at MIT CSAIL advised by Jacob Andreas and Julie Shah . I've spent summers at Facebook AI Research (FAIR), Boston Dynamics AI Institute, MIT-IBM Watson AI Lab, and before grad school, two years as an AI Resident at Microsoft Research. I did my undergrad at Yale, where I got my start in research with Brian Scassellati and read a lot of dead philosophers. I'm interested in building agents that learn representations from
Explore this link on the map →saved by
related reading
- Group | Sherry Tongshuang Wucs.cmu.edu
- Vibe physics: The AI grad student \ Anthropicanthropic.com
- Automated Weak-to-Strong Researcheralignment.anthropic.com
- Andrej Karpathy — AGI is still a decade awaydwarkesh.com
- AI in 2025: gestalt — LessWronglesswrong.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Automated Alignment Researchers: Using large language models to scale scalable oversight \ Anthropicanthropic.com
- RLHF: Reinforcement Learning from Human Feedbackhuyenchip.com
- pdfopenreview.net
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Teaching Claude why \ Anthropicanthropic.com
- [AN #70]: Agents that help humans who are still learning about their own preferences — LessWronglesswrong.com