Simulators - LessWrong
This post explores the concept of simulators in AI, particularly self-supervised models like GPT. Janus argues that GPT and similar models are best u…
x Simulators — LessWrong Simulators Simulator Theory Language Models (LLMs) LLM Personas GPT Outer Alignment Simulation Corrigibility Deconfusion Myopia Oracle AI Tool AI AI Curated 713 Simulators by janus 2nd Sep 2022 AI Alignment Forum Linkpost for generative.ink 49 min read 170 713 Ω 144 Thanks to Chris Scammell, Adam Shimi, Lee Sharkey, Evan Hubinger, Nicholas Dupuis, Leo Gao, Johannes Treutlein, and Jonathan Low for feedback on drafts. This work was carried out while at Conjecture . "Moebius illustration of a simulacrum living in an AI-generated story discovering it is in a simulation" by
Explore this link on the map →saved by
- Alex K. Chen
- Bryan Chiang
- Madison Ueland
- Micah Carroll
- Kevin Wang
- Karan MJ
- Varun Shenoy
- Ratan Kaliani
- Atem Aguer
- Arden Berg
- Lydia Nottingham
- Yudhister Joel Kumar
related reading
- Simulators :: — Moiregenerative.ink
- trees are harlequins, words are harlequins - the voidnostalgebraist.tumblr.com
- Cyborgism — LessWronglesswrong.com
- The Persona Selection Model: Why AI Assistants might Behave like Humansalignment.anthropic.com
- Why Simulator AIs want to be Active Inference AIs — AI Alignment Forumalignmentforum.org
- Role-playing vs Self-modelling — LessWronglesswrong.com
- gpt-4.pdfcdn.openai.com
- Commentary On The Turing Apocryphaminihf.com
- GPTs are Predictors, not Imitators — AI Alignment Forumalignmentforum.org
- The Waluigi Effect (mega-post) — LessWronglesswrong.com
- Cyborgism — AI Alignment Forumalignmentforum.org
- Andrej Karpathy — AGI is still a decade awaydwarkesh.com