You know what the hardware is. Now: how do you tell it what to do?
This module teaches the CUDA programming model — the one nearly every other GPU API copies — and shows how a Go program drives it.
| # | Lesson | The question it answers |
|---|---|---|
| 01 | Kernels, Threads, Blocks, Grids | How is work described to a GPU? |
| 02 | Your First Kernel, Called from Go | What does a complete GPU program look like? |
| 03 | Host and Device Memory | Where does data live, and how does it move? |
| 04 | Asynchronous Execution and Streams | Why does a GPU call return before the work is done? |
| 05 | The Software Stack | What sits between go run and the silicon? |
Lesson 02 needs an NVIDIA GPU and the CUDA toolkit to run; everything else runs anywhere. If you have no GPU, read 02 carefully and do its simulator exercise instead — the concepts carry.