✳flâneur — a map of the web's best reading
[Jan 7 2026] nanochat miniseries v1 · karpathy/nanochat · Discussion #420
github.com · 7,915 words · saved by 1 readers
Why miniseries. The correct way to think about LLMs is that you are not optimizing for a single specific model but for a family models controlled by a single dial (the compute you wish to spend) to...
Explore this link on the map →saved by
related reading
- LLM Visualizationbbycroft.net
- GitHub - karpathy/nanochat: The best ChatGPT that $100 can buy. · GitHubgithub.com
- LLM Daydreaming · Gwern.netgwern.net
- GenAI Handbookgenai-handbook.github.io
- GitHub - jacobhilton/deep_learning_curriculum: Language model alignment-focused deep learning curriculum · GitHubgithub.com
- GitHub - openai/parameter-golf: Train the smallest LM you can that fits in 16MB. Best model wins! · GitHubgithub.com
- GitHub - karpathy/autoresearch: AI agents running research on single-GPU nanochat training automatically · GitHubgithub.com
- MatX: High-throughput chips for LLMsmatx.com
- 2025: The year in LLMssimonwillison.net
- Things we learned about LLMs in 2024simonwillison.net
- Model optimization | OpenAI APIplatform.openai.com
- Direct Nash Optimization: Teaching Language Models to Self-Improve with General Preferencesarxiv.org