✳flâneur — a map of the web's best reading
100M Token Context Windows — Magic
magic.dev · 1,801 words · saved by 1 readers
Research update on ultra-long context models, our partnership with Google Cloud, and new funding.
100M Token Context Windows Research update on ultra-long context models, our partnership with Google Cloud, and new funding. Magic Team , on August 29, 2024 There are currently two ways for AI models to learn things: training, and in-context during inference. Until now, training has dominated, because contexts are relatively short. But ultra-long context could change that. Instead of relying on fuzzy memorization, our LTM (Long-Term Memory) models are trained to reason on up to 100M tokens of context given to them during inference. While the commercial applications of these ultra-long context
Explore this link on the map →related reading
- Recursive Language Models | Alex L. Zhangalexzhang13.github.io
- Composer2.pdfcursor.com
- My picture of the present in AI — LessWronglesswrong.com
- Reimagining LLM Memory: Using Context as Training Data Unlocks Models That Learn at Test-Time | NVIDIA Technical Blogdeveloper.nvidia.com
- The Short Case for Nvidia Stock | YouTube Transcript Optimizeryoutubetranscriptoptimizer.com
- What is a context window? | IBMibm.com
- gemini_v1_5_report.pdfstorage.googleapis.com
- Extending Context is Hard | kaiokendevkaiokendev.github.io
- [2506.06266] Cartridges: Lightweight and general-purpose long context representations via self-studyarxiv.org
- Mediumblog.gopenai.com
- Generalizing an LLM from 8k to 1M Context using Qwen-Agent | Qwenqwenlm.github.io
- Subquadratic — How SSA Makes Long Context Practicalsubq.ai