3 ms·
Yeah I’m honestly unclear on Nvidia’s thinking here - inference speed is unbelievably slow for the price. Given the extreme advantage they have with CUDA and t
by yunohn 1y ago
Yeah I’m honestly unclear on Nvidia’s thinking here - inference speed is unbelievably slow for the price.
Given the extreme advantage they have with CUDA and the whole AI/ML ecosystem, barely matching Apple’s M-ultra speeds is a choice…
- airspresso 1y agoDefinitely a choice to give it low memory bandwidth. Probably to avoid customers thinking it can replace any data center GPU for inference use-cases.