Simulators - LessWrong
This post explores the concept of simulators in AI, particularly self-supervised models like GPT. Janus argues that GPT and similar models are best u…
x Simulators — LessWrong Simulators Simulator Theory Language Models (LLMs) LLM Personas GPT Outer Alignment Simulation Corrigibility Deconfusion Myopia Oracle AI Tool AI AI Curated 713 Simulators by janus 2nd Sep 2022 AI Alignment Forum Linkpost for generative.ink 49 min read 170 713 Ω 144 Thanks to Chris Scammell, Adam Shimi, Lee Sharkey, Evan Hubinger, Nicholas Dupuis, Leo Gao, Johannes Treutlein, and Jonathan Low for feedback on drafts. This work was carried out while at Conjecture . "Moebius illustration of a simulacrum living in an AI-generated story discovering it is in a simulation" by
saved by
- Alex K. Chen
- Bryan Chiang
- Madison Ueland
- Micah Carroll
- Kevin Wang
- Karan MJ
- Varun Shenoy
- Ratan Kaliani
- Atem Aguer
- Arden Berg
- Lydia Nottingham
- Gene Yang
related reading
- Simulators :: — Moiregenerative.ink
- Janus' Simulatorsastralcodexten.substack.com
- trees are harlequins, words are harlequins - the voidnostalgebraist.tumblr.com
- Cyborgism — LessWronglesswrong.com
- Why Simulator AIs want to be Active Inference AIs — AI Alignment Forumalignmentforum.org
- The Persona Selection Model: Why AI Assistants might Behave like Humansalignment.anthropic.com
- gpt-4.pdfcdn.openai.com
- Role-playing vs Self-modelling — LessWronglesswrong.com
- GPTs are Predictors, not Imitators — AI Alignment Forumalignmentforum.org
- Commentary On The Turing Apocryphaminihf.com
- The Waluigi Effect (mega-post) — LessWronglesswrong.com
- Cyborgism — AI Alignment Forumalignmentforum.org