3 ms·
I thought these TPUs were primarily used for inference?
by written-beyond 6mo ago
I thought these TPUs were primarily used for inference?
- vlovich123 6mo agoTPU8t is for training. But even still, once you’ve trained, you need to run the model too. And these kinds of models already have a huge latency hit so there’s not much hurting running it away from the trading switches.
- knowaveragejoe 6mo agoAs the article states, there's both training and inference dedicated chips.