Subquadratic — Efficiency is Intelligence
subq.ai · 2,497 words · saved by 2 readers
Subquadratic is a frontier AI research and infrastructure company building a new class of LLMs.
Subquadratic — How SSA Makes Long Context Practical Research How SSA Makes Long Context Practical Date May 5, 2026 Updated May 15, 2026 Updated with recent third-party validated benchmarks from source Note: In this paper we share third-party verified benchmarks. A comprehensive model card is coming soon! SubQ is built around SSA, Subquadratic Sparse Attention, a linearly scaling attention mechanism designed for long-context retrieval, reasoning, and software engineering workloads. The core claim is simple: the hard problems enterprise AI needs to solve are long-context problems. Codebase, cont
saved by
related reading
- Subquadratic — Introducing SubQ: The First Fully Subquadratic LLMsubq.ai
- [2507.04239] Scaling Context Requires Rethinking Attentionarxiv.org
- 2502.11089arxiv.org
- A short note on some aspects of long context attention | nor's blognor-blog.pages.dev
- On the Tradeoffs of SSMs and Transformers | Goomba Labgoombalab.github.io
- [2603.23516] MSA: Memory Sparse Attention for Efficient End-to-End Memory Model Scaling to 100M Tokensarxiv.org
- The Big LLM Architecture Comparisonmagazine.sebastianraschka.com
- [2604.20920] Simplified Sparse Attention via Gist Tokensarxiv.org
- Quest: Query-Aware Sparsity for Efficient Long-Context LLM Inferencearxiv.org
- Linear Attention Fundamentals | Hailey Schoelkopfhaileyschoelkopf.github.io
- Kimi Linear: An Expressive, Efficient Attention Architecturearxiv.org
- DeltaNet Explained (Part I) | Songlin Yangsustcsonglin.github.io