The social AI hypothesis - by Stefano Viel - Ste
substack.com · 1,528 words · saved by 2 readers
Training a neural network induces a policy that performs well under a particular objective and data distribution.
The current approach to improving LLMs’ capabilities is to create a sufficiently diverse set of environments, train on all of them, and hope to obtain generality. This requires continuous human contribution and, unlike pre-training, doesn’t scale. I believe that training in an environment with other agents (i.e., social training) scales with the number of agents in the environment as they create additional complexity and require more intelligence without any human intervention. Thus, social training represents a more viable path to artificial general intelligence. Don’t get me wrong, what…
saved by
related reading
- Will scaling work?dwarkeshpatel.com
- On neural scaling and the quanta hypothesisericjmichaud.com
- AI 2027ai-2027.com
- Ilya Sutskever — We're moving from the age of scaling to the age of researchdwarkesh.com
- Andrej Karpathy — AGI is still a decade awaydwarkesh.com
- The Scaling Hypothesis · Gwern.netgwern.net
- Questions about the Future of AI - by Dwarkesh Pateldwarkesh.com
- Will scaling work? - by Dwarkesh Patel - Dwarkesh Podcastdwarkesh.com
- Just Ask for Generalization | Eric Jangevjang.com
- The upcoming GPT-3 moment for RL | Mechanize, Inc.mechanize.work
- AI in 2025: gestalt — LessWronglesswrong.com
- Three Observations - Sam Altmanblog.samaltman.com