Modular: What exactly is “CUDA”? (Democratizing AI Compute, Part 2)
It seems like everyone has started talking about CUDA in the last year: It’s the backbone of deep learning, the reason novel hardware struggles to compete, and the core of NVIDIA’s moat and soaring market cap. With DeepSeek, we got a startling revelation: its breakthrough was made possible by “bypassing” CUDA, going directly to the PTX layer… but what does this actually mean? It feels like everyone wants to break past the lock-in, but we have to understand what we’re up against before we can formulate a plan. This is Part 2 of Modular’s “Democratizing AI Compute” series. For more, see: CUDA’s dominance in AI is undeniable—but most people don’t fully understand what CUDA actually is. Some think it’s a programming language. Others call it a framework. Many assume it’s just “that thing NVIDIA uses to make GPUs faster.” None of these are entirely wrong—and many brilliant people are trying to explain this—but none capture the full scope of “The CUDA Platform.” CUDA is not just one thing. It
Modular: What exactly is “CUDA”? (Democratizing AI Compute, Part 2) Qualcomm to Acquire Modular. Read More → February 5, 2025 What exactly is “CUDA”? (Democratizing AI Compute, Part 2) Chris Lattner Series It seems like everyone has started talking about CUDA in the last year: It’s the backbone of deep learning, the reason novel hardware struggles to compete, and the core of NVIDIA’s moat and soaring market cap. With DeepSeek, we got a startling revelation: its breakthrough was made possible by “bypassing” CUDA , going directly to the PTX layer … but what does this actually mean? It feels like
Explore this link on the map →related reading
- The Short Case for Nvidia Stock | YouTube Transcript Optimizeryoutubetranscriptoptimizer.com
- CUDA - Wikipediaen.wikipedia.org
- CUDA C++ Programming Guide (Legacy) — CUDA C++ Programming Guidedocs.nvidia.com
- Making Deep Learning go Brrrr From First Principleshorace.io
- The Best GPUs for Deep Learning in 2023 — An In-depth Analysistimdettmers.com
- Nvidia Part II: The Machine Learning Company (2006-2022) | Acquiredacquired.fm
- A friendly introduction to machine learning compilers and optimizershuyenchip.com
- How Nvidia’s CUDA Monopoly In Machine Learning Is Breaking - OpenAI Triton And PyTorch 2.0semianalysis.com
- An Even Easier Introduction to CUDA (Updated) | NVIDIA Technical Blogdeveloper.nvidia.com
- How to Think About GPUs | How To Scale Your Modeljax-ml.github.io
- Modular: Inference from Kernel to Cloudmodular.com
- [2410.20399] ThunderKittens: Simple, Fast, and Adorable AI Kernelsarxiv.org