What is a Streaming Multiprocessor? | GPU Glossary
When we program GPUs , we produce sequences of instructions for its Streaming Multiprocessors to carry out. A diagram of the internal architecture of an H100 GPU's Streaming Multiprocessors. GPU cores appear in green, other compute units in maroon, scheduling units in orange, and memory in blue. Modified from NVIDIA's H100 white paper . A diagram of the internal architecture of an H100 GPU's Streaming Multiprocessors. GPU cores appear in green, other compute units in maroon, scheduling units in orange, and memory in blue. Modified from NVIDIA's H100 white paper . Streaming Multiprocessors (SMs) of NVIDIA GPUs are roughly analogous to the cores of CPUs. That is, SMs both execute computations and store state available for computation in registers, with associated caches. Compared to CPU cores, GPU SMs are simple, weak processors. Execution in SMs is pipelined within an instruction (as in almost all CPUs since the 1990s) but there is no speculative execution or instruction pointer predic
What is a Streaming Multiprocessor? | GPU Glossary GPU Glossary GPU Glossary Terminal Light green Light Deploy on GPUs TABLE OF CONTENTS Home - README Device Hardware - CUDA (Device Architecture) Streaming Multiprocessor SM Core Special Function Unit SFU Load/Store Unit LSU Warp Scheduler CUDA Core Tensor Core Tensor Memory Accelerator TMA Streaming Multiprocessor Architecture Texture Processing Cluster TPC Graphics/GPU Processing Cluster GPC Register File L1 Data Cache Tensor Memory GPU RAM Device Software - CUDA (Programming Model) Streaming ASSembler SASS Parallel Thread eXecution PTX Compu
Explore this link on the map →related reading
- Execution Model - SLING user documentationdoc.sling.si
- Inside NVIDIA GPUs: Anatomy of high performance matmul kernels - Aleksa Gordićaleksagordic.com
- A history of NVidia Stream Multiprocessorfabiensanglard.net
- Demystifying GPU Compute Architectures - by Babbagethechipletter.substack.com
- CUDA C++ Programming Guide (Legacy) — CUDA C++ Programming Guidedocs.nvidia.com
- How to Think About GPUs | How To Scale Your Modeljax-ml.github.io
- Gentle introduction to GPUs inner workings | vkSegfaultvksegfault.github.io
- BrrrVizbrrrviz.com
- The Best GPUs for Deep Learning in 2023 — An In-depth Analysistimdettmers.com
- How to Optimize a CUDA Matmul Kernel for cuBLAS-like Performance: a Worklogsiboehm.com
- GPU Performance Background User's Guide - NVIDIA Docsdocs.nvidia.com
- General-purpose computing on graphics processing units - Wikipediaen.wikipedia.org