Kimi Linear: An Expressive, Efficient Attention Architecture
arxiv.org · 5,791 words · saved by 2 readers
N/A
K IMI L INEAR : A N E XPRESSIVE , E FFICIENT ATTENTION A RCHITECTURE T ECHNICAL R EPORT OF K IMI L INEAR Kimi Team https://github.com/MoonshotAI/Kimi-Linear arXiv:2510.26692v2 [cs.CL] 1 Nov 2025…
saved by
related reading
- Kimi Linear: An Expressive, Efficient Attention Architecturealphaxiv.org
- [2510.26692] Kimi Linear: An Expressive, Efficient Attention Architecturearxiv.org
- ali (@waterloo_intern) on Xx.com
- The Big LLM Architecture Comparisonmagazine.sebastianraschka.com
- Linear Attention Fundamentals | Hailey Schoelkopfhaileyschoelkopf.github.io
- DeltaNet Explained (Part I) | Songlin Yangsustcsonglin.github.io
- [2607.07953] Linear Attention Architectures: Mechanisms, Trade-offs, and Cross-Layer Routingarxiv.org
- Log-Linear Attentionarxiv.org
- Linear Transformers Are Faster After All – Manifest AImanifestai.com
- 2502.11089arxiv.org
- Mamba: The Easy Wayjackcook.com
- Transformers Inference Optimization Toolset | AstraBlogastralord.github.io