4 ms·
Just recently read the interview transcript [0] of CEO and CTO (Jim Keller) of Tenstorrent which I believe they're in the same field. That's interestingly deep
by wejick 5y ago
Just recently read the interview transcript [0] of CEO and CTO (Jim Keller) of Tenstorrent which I believe they're in the same field. That's interestingly deep conversation open my eye a little bit about AI processor and how it's not only about a chip that accelerate "AI task".
[0] https://www.anandtech.com/show/16709/an-interview-with-tenstorrent-ceo-ljubisa-bajic-and-cto-jim-keller https://www.anandtech.com/show/16709/an-interview-with-tenst...
- dragontamer 5y agoThese "AI processors" are just matrix-multiplication engines, often 8-bit or 16-bits. Maybe 4-bit in some cases. 16-bit (and smaller) just isn't enough for most compute problems. But emphasis on "most". Neural Nets clearly are fine with 16-bit, but some video game lighting effects can be calculated on 16-bit and look "good enough". Maybe a certified Jewelry Appraiser can tell the difference between 1.3 refractive index vs 1.35, but the typical video game player (and dare I say, human) won't be able to tell the difference. If all the light bounces are just slightly off, as long as its "close enough", you probably have good-enough looking ray-tracing or whatever. -------- I've also heard of "iterative solvers" being accelerated by 16-bit and 32-bit passes before doing a 64-bit pass. That's not applicable to all math problems of course, but that's still a methodology where your 16-bit performance accelerates your 64-bit ultimate answer.
- lostmsu 5y agoFor casual readers, the trend in neural nets has changed a bit, and for training TensorFloat-32 is gaining popularity: https://blogs.nvidia.com/blog/2020/05/14/tensorfloat-32-precision-format/ https://blogs.nvidia.com/blog/2020/05/14/tensorfloat-32-prec... Pure f16 overflows/underflows in many scenarios