4 ms·
AFAIK, they don't have any deals or partnerships with Groq or Cerebras or any of those kinds of companies.. so how did they do this?
by maz1b 8mo ago
AFAIK, they don't have any deals or partnerships with Groq or Cerebras or any of those kinds of companies.. so how did they do this?
- hendersoon 8mo agoCould well be running on Google TPUs.
- tcdent 8mo agoInference is run on shared hardware already, so they're not giving you the full bandwidth of the system by default. This most likely just allocates more resources to your request.
- deleted 8mo ago[deleted]
- rvz 8mo agoThe models are running on Google TPUs.