3 ms·
The AnandTech article on the Volta has a lot more information on the new architecture: http://www.anandtech.com/show/11367/nvidia-volta-unveiled-gv100-gpu-and-t
by DocSavage 9y ago
The AnandTech article on the Volta has a lot more information on the new architecture:
http://www.anandtech.com/show/11367/nvidia-volta-unveiled-gv100-gpu-and-tesla-v100-accelerator-announced http://www.anandtech.com/show/11367/nvidia-volta-unveiled-gv...
It's interesting the speed up isn't more pronounced between Volta and Pascal considering the Tensor cores on paper give you about 6x the MFlops. The price differential looks large.
From AnandTech: "By the numbers, Tesla V100 is slated to provide 15 TFLOPS of FP32 performance, 30 TFLOPS FP16, 7.5 TFLOPS FP64, and a whopping 120 TFLOPS of dedicated Tensor operations. With a peak clockspeed of 1455MHz, this marks a 42% increase in theoretical FLOPS for the CUDA cores at all size. Whereas coming from Pascal, for Tensor operations the gains will be closer to 6-12x, depending on the operation precision."
- trueSlav 9y agoNo free lunch and all that... 40% seems quite nice if you think about the transistor count (15b to 21b).