4 ms·
This is pretty awesome.
by hellofunk 7y ago
This is pretty awesome.
- sv_h1b 7y agoGPUs are designed to do massive parallel computations which all branch one way and don't have much (any?) branch prediction logic. You trade it off with increased code size which might spill the cache but for small tight loops not exceeding the cache line/ size it would still be a good win.
- ScottFree 7y agoWould there be any benefit to adding branch prediction logic to a GPU?
- rdc12 7y agoYou would probably get some benefit on a per CUDA core basis, but the extra space required would mean less CUDA cores, overall it would probably be a net loss on problems well suited to GPU acceleration