flâneur — a map of the web's best reading

chinchilla's wild implications — AI Alignment Forum

alignmentforum.org · 5,860 words · saved by 1 readers

(Colab notebook here.) • This post is about language model scaling laws, specifically the laws derived in the DeepMind paper that introduced Chinchilla.[1] …

x chinchilla's wild implications — AI Alignment Forum Best of LessWrong 2022 Scaling Laws Language Models (LLMs) Machine Learning (ML) DeepMind AI Frontpage 97 chinchilla's wild implications by nostalgebraist 31st Jul 2022 13 min read 129 97 ( Colab notebook here.) This post is about language model scaling laws, specifically the laws derived in the DeepMind paper that introduced Chinchilla. [1] The paper came out a few months ago, and has been discussed a lot, but some of its implications deserve more explicit notice in my opinion. In particular: Data, not size, is the currently active constra

Explore this link on the map →

related reading