parallel-computing-overview
Differences
This shows you the differences between two versions of the page.
| Both sides previous revisionPrevious revision | |||
| parallel-computing-overview [June 15, 2026 at 09:34] – Ivan Janevski | parallel-computing-overview [August 22, 2026 at 15:22] (current) – external edit 127.0.0.1 | ||
|---|---|---|---|
| Line 43: | Line 43: | ||
| On the performance side, the [[roofline-model|roofline model]] is a useful frame for understanding whether a kernel is compute-bound or memory-bandwidth-bound, | On the performance side, the [[roofline-model|roofline model]] is a useful frame for understanding whether a kernel is compute-bound or memory-bandwidth-bound, | ||
| - | ## Practice | + | ## Example |
| The fastest way to see parallelism pay off is to compile the same program twice — once without threading, once with — and time both. Here is a parallel sum using OpenMP: | The fastest way to see parallelism pay off is to compile the same program twice — once without threading, once with — and time both. Here is a parallel sum using OpenMP: | ||
parallel-computing-overview.1781516094.md.gz · Last modified: by Ivan Janevski
