Lecture
Mediaspace scheduled maintenance: Aug 25, 2026 07:00 - 12:00 AM. During this time, videos will be temporarily unavailable. Check status updates.
This lecture covers the GPU memory hierarchy, including global, local, shared memory, and caches, and discusses the challenges of SIMT execution. It delves into optimizing algorithms for GPUs by coalescing accesses, reducing bank conflicts, and eliminating warp divergence. The lecture also emphasizes the importance of understanding the algorithm's nature to optimize memory-intensive code efficiently.
This video is available exclusively on Mediaspace for a restricted audience. Please log in to MediaSpace to access it if you have the necessary permissions.
Watch on Mediaspace