flâneur — a map of the web's best reading

Expanding AI Control from Models to Harnesses — LessWrong

lesswrong.com · 6,033 words · saved by 1 readers

Most AI control research such as LinuxArena and Ctrl-Z only gives the red team basic agents which only have access to tools. Yet in 2026, usage of AI…

x Expanding AI Control from Models to Harnesses — LessWrong AI "Agent" Scaffolds AI Control AI Frontpage 17 Expanding AI Control from Models to Harnesses by fastfedora 15th Jul 2026 24 min read 0 17 Most AI control research such as LinuxArena and Ctrl-Z only gives the red team basic agents which only have access to tools. Yet in 2026, usage of AI within frontier labs has moved to agent harnesses that have access to skills, memory, subagents, external services, compaction and more. At the same time, Claude Code and Codex have both implemented their own version of both action-based and source co

Explore this link on the map →

related reading