5 ms·
The GPUs that hosting providers are buying need to support vGPU features, which are only support by the higher-end workstation and server products.
by perryh2 10y ago
The GPUs that hosting providers are buying need to support vGPU features, which are only support by the higher-end workstation and server products.
- LogicFailsMe 10y agoWhich is a software/binning issue (probably 100% software) not inherently HW because M60==GTX 980==GM204 and M40==GTX Titan X==Quadro M6000==GM200. Interestingly, for the first time ever, GP100 is unique (well, OK, K80 too but K80 was too late). And since Quadro P6000==Titan XP==GP102, it's probably just a software block here as well. Also, for the first time ever, the high-end Quadro will be the best FP32 GPU of it's generation. That's the most interesting part for me.
- dogma1138 10y agoIt is a software/bios block. Once the driver grabs a desktop card and it's initiated it's locked. The cards also come with a UEFI bios which means if your host initialized the card during boot it's also locked. To overcome this you need to disable UEFI boot/GPU boot in the BIOS and blacklist the card in the host OS and then create a PCI-stub device that will be used for the passthrough. This is an utter and complete hack and you can't really use for production grade implementations. I'm not sure that Quadro actually supports vGPU also AFAIK only Tesla and Grid parts do, Tesla does have some additional in GPU support for virtualization, i don't know how GRID handles it. GRID parts are less for compute and more for game/video streaming so they aren't fully virtualized, AFAIK they do not support P2P gpu communication or shared memory access, they are only more Hypervisor friendly so they can be thin provisioned.
- LogicFailsMe 10y agoBare metal grid GPUs do support P2P. Xen breaks that on AWS.