AI has entered the chat - by Ben Lee and Paulius Mui, MD
In 2017, the Google Brain team’s landmark paper Attention Is All You Need introduced the Transformer architecture, which drastically improved performance and quality in building large language models (LLMs). What followed was a convergence in approach across ML / AI research, resulting in composability of ideas and compounding of improvements. 2020 saw the OpenAI team publish Scaling Laws for Neural Language Models which led to the emergence of the Scaling Hypothesis - TL;DR: pump model size, data, and compute and you’ll see increasingly sophisticated behavior in AI models. Then of course, a small “research preview” called ChatGPT surfaced the capabilities and potential of LLMs through a consumer-facing interface, and became the fastest growing consumer app of all time (100M MAUs in just two months!). Here’s a fantastic primer on most of the concepts behind the tech. The resulting Cambrian explosion of AI innovation has occurred at a breakneck pace - so much so that, a few days ago, ma
In 2017, the Google Brain team’s landmark paper Attention Is All You Need introduced the Transformer architecture, which drastically improved performance and quality in building large language models (LLMs). What followed was a convergence in approach across ML / AI research, resulting in composability of ideas and compounding of improvements. 2020 saw the OpenAI team publish Scaling Laws for Neural Language Models which led to the emergence of the Scaling Hypothesis - TL;DR: pump model size, data, and compute and you’ll see increasingly sophisticated behavior in AI models. Then of course, a s
Explore this link on the map →