flâneur — a map of the web's best reading

A basic systems architecture for AI agents that do autonomous research — LessWrong

lesswrong.com · 5,129 words · saved by 4 readers

A lot of threat models describing how AIs might escape our control (e.g. self-exfiltration, hacking the datacenter) start out with AIs that are actin…

x A basic systems architecture for AI agents that do autonomous research — LessWrong Redwood Research AI Curated 190 A basic systems architecture for AI agents that do autonomous research by Buck 23rd Sep 2024 AI Alignment Forum 9 min read 17 190 Ω 71 A lot of threat models describing how AIs might escape our control (e.g. self-exfiltration , hacking the datacenter ) start out with AIs that are acting as agents working autonomously on research tasks (especially AI R&D) in a datacenter controlled by the AI company. So I think it’s important to have a clear picture of how this kind of AI agent c

Explore this link on the map →

saved by

related reading