✳flâneur — a map of the web's best reading
GPT-4 Architecture, Infrastructure, Training Dataset, Costs, Vision, MoE
semianalysis.com · 1,267 words · saved by 1 readers
Demystifying GPT-4: The engineering tradeoffs that led OpenAI to their architecture.
GPT-4 Architecture, Infrastructure, Training Dataset, Costs, Vision, MoE Demystifying GPT-4: The engineering tradeoffs that led OpenAI to their architecture. Dylan Patel and Gerald Wong Jul 10, 2023 ∙ Paid 197 50 Share OpenAI is keeping the architecture of GPT-4 closed not because of some existential risk to humanity but because what they’ve built is replicable. In fact, we expect Google, Meta, Anthropic, Inflection, Character, Tencent, ByteDance, Baidu, and more to all have models as capable as GPT-4 if not more capable in the near term. Don’t get us wrong, OpenAI has amazing engineering, and
Explore this link on the map →saved by
related reading
- How To Scale Your Modeljax-ml.github.io
- gpt-4.pdfcdn.openai.com
- The Scaling Hypothesis · Gwern.netgwern.net
- Google "We Have No Moat, And Neither Does OpenAI"semianalysis.com
- Reiner Pope – The math behind how LLMs are trained and serveddwarkesh.com
- GPT-4openai.com
- I. From GPT-4 to AGI: Counting the OOMs - SITUATIONAL AWARENESSsituational-awareness.ai
- My picture of the present in AI — LessWronglesswrong.com
- The Short Case for Nvidia Stock | YouTube Transcript Optimizeryoutubetranscriptoptimizer.com
- What will GPT-2030 look like? — AI Alignment Forumalignmentforum.org
- LLM Inference Performance Engineering: Best Practices | Databricks Blogdatabricks.com
- Navigating the High Cost of AI Compute | Andreessen Horowitza16z.com