Mediaspace scheduled maintenance: Aug 25, 2026 07:00 - 12:00 AM. During this time, videos will be temporarily unavailable. Check status updates.
This lecture by the instructor covers software optimizations focusing on locality, memory access, and scheduling strategies for parallel execution. It delves into cache hierarchy, latency considerations, cache miss models, coherence misses, and techniques to reduce true and false sharing. The lecture also includes examples like histogram computation, parallel work division, and matrix multiplication to illustrate optimization strategies. It emphasizes the importance of the locality principle, blocking for cache efficiency, and load balancing through dynamic work distribution. Additionally, it discusses loop optimizations, task queues for parallel processing, and functional parallelism for independent tasks.
This video is available exclusively on Mediaspace for a restricted audience. Please log in to MediaSpace to access it if you have the necessary permissions.
Watch on Mediaspace