NVIDIA CUDA Tutorial 10: Blocking with Shared Memory

Pubblicato il: 16 dicembre 2013
sul canale di: Creel
20,408
211

In this tute we'll use a technique called blocking to finally fulfill Porky Water's tall order!

Blocking is a technique where blocks of data are copied from global memory to shared memory, threads work on the data in the much faster shared memory. This greatly reduces the amount of traffic on the global memory bus and allows threads to use the much faster shared memory for most of the calculations.

Blocking with shared memory gives us a great speed up here and easily fulfills Porky's boss's request of a 10x speed up. There's some small changes that could allow the code to run a little quicker but if the code had to run much faster a complete change in algorithm would be far more useful than tweaking this brute force one.


In questa pagina del sito puoi guardare il video online NVIDIA CUDA Tutorial 10: Blocking with Shared Memory della durata di ore minuti seconda in buona qualità , che l'utente ha caricato Creel 16 dicembre 2013, condividi il link con amici e conoscenti, su youtube questo video è già stato visto 20,408 volte e gli è piaciuto 211 spettatori. Buona visione!