MODULE 06 / Days 51-60

Concurrency

All lessons
  1. 51
    CUDA streams and overlapHow do CUDA streams overlap two kernels?
    run logged
  2. 52
    Events and cross-stream dependenciesHow do I make one CUDA stream wait on another?
    run logged
  3. 53
    Pinned memory and async copiesWhy is cudaMemcpyAsync not asynchronous?
    run logged
  4. 54
    Double bufferingHow does double buffering overlap CUDA copies and kernels?
    run logged
  5. 55
    Stream-ordered allocation and memory poolsWhat does cudaMallocAsync actually do?
    run logged
  6. 56
    CUDA graphsWhen do CUDA graphs actually make a program faster?
    run logged
  7. 57
    Graph updates and conditional nodesHow do I change a CUDA graph's parameters, and can a graph loop until convergence without the host?
    run logged
  8. 58
    Programmatic dependent launchWhat is programmatic dependent launch in CUDA?
    lesson
  9. 59
    Host threads and the GPUCan multiple CPU threads use CUDA at the same time?
    run logged
  10. 60
    Capstone 3: a real-time frame pipelineHow do I build a real-time video pipeline in CUDA?
    run logged