3 ms·
Look into groq.com guys. some good models at similar speed to inception labs
by ZeroTalent 1y ago
Look into groq.com guys. some good models at similar speed to inception labs
- sujayk_33 1y agoIt's faster inference because of the Hardware (LPUs), here the question is about architectures (AR or Diffusions)
- ZeroTalent 1y agoI realize that, but it can be used now with many models in real-life situations. I just wanted to mention it if someone doesn't know it.
- rfv6723 1y agoSRAM doesn't scale with advanced semiconductor node. Groq is heading to a dead end.