About Me | Amanda Askell
I’m a philosopher working on AI alignment at Anthropic. Before this I worked as a research scientist on the policy team at OpenAI, where I worked on AI safety via debate and human baselines for AI performance. I have a PhD in philosophy from NYU with a thesis on infinite ethics and a BPhil in philosophy from the University of Oxford. My philosophy work is mostly in ethics, decision theory, and formal epistemology. I’ve been involved in effective altruism since around 2010. I’m a member of Giving What We Can and have appeared on Rationally Speaking and the 80,000 hours podcast.
I’m a philosopher working on finetuning and AI alignment at Anthropic. My team trains models to be more honest and to have good character traits, and works on developing new finetuning techniques so that our interventions can scale to more capable models. Before this I worked as a research scientist on the policy team at OpenAI, where I worked on AI safety via debate and human baselines for AI performance. I have a PhD in philosophy from NYU with a thesis on infinite ethics and a BPhil in philosophy from the University of Oxford. I did my undergraduate degree in Philosophy at the University…
saved by
related reading
- Leaving Open Philanthropy, going to Anthropic - Joe Carlsmithjoecarlsmith.com
- Amanda Askellen.wikipedia.org
- LessWronglesswrong.com
- guzey2guzey2.com
- About Gavin Leechgleech.org
- The Best of LessWrong — LessWronglesswrong.com
- ALIGNMENT - by vincent huang - a slice of my mindmindslice.substack.com
- Home - Joe Carlsmithjoecarlsmith.com
- [About Me] Cinera's Home Page — LessWronglesswrong.com
- ALIGNMENT - by vincent huang - a slice of my mindmindslice.substack.com
- Anna Wangannaywang.com
- The Universe from an Intentional Stancecasparoesterheld.com