Module I said a GPU has “thousands of lanes” and “fast memory”. This module opens the lid and names every part, using NVIDIA’s H100 as the worked example.
| # | Lesson | The question it answers |
|---|---|---|
| 01 | SMs, Warps and SIMT | How are thousands of lanes organized and controlled? |
| 02 | The Memory Hierarchy | Where can data live on a GPU, and how fast is each place? |
| 03 | HBM and Memory Bandwidth | What is the memory made of, and why is its speed the headline? |
| 04 | Tensor Cores and Number Formats | Where do the giant FLOP numbers come from? |
| 05 | Power, Heat and the Host Link | What connects the chip to the rest of the world? |
When you finish you can draw a GPU from memory — SMs, warps, registers, shared memory, caches, HBM, PCIe — and attach a rough speed and size to every box.