Moving data from memory to GPU compared to how we do in processor

 The Comparison: Microcontroller vs. GPU

To make this crystal clear, let's look at how a standard CPU/Microcontroller compiler thinks, versus how an AI GPU compiler thinks when performing the exact same task: Adding two giant arrays of numbers.
ConceptMicrocontroller / CPU Compiler ViewAI GPU Compiler View
Where is the Program Stored?In flash memory / EEPROM / System RAM.In VRAM (loaded into a specific segment of the GPU's memory).
How Instructions are FetchedOne by one via the Program Counter (PC) into the instruction register.One instruction is fetched by a control unit, but it is broadcasted to thousands of cores at once.
How Registers are LoadedMVI A, 09
MOV B, M
(Loading one tiny register at a time).
LOAD_MATRIX R0, [VRAM_ADDR_A]
(Thousands of different data slots load into thousands of different core registers simultaneously).
The Math ExecutionADD A, B
(Executed once for one pair of numbers).
MATRIX_MUL R2, R1, R0
(Executed once, but thousands of internal ALUs fire at the exact same millisecond).

Comments

Popular posts from this blog

how to add all current and future projects of android studio to allow in windows firewall security..

adjusting width of explorrer in xamp project folder file names - php LocalHost