4 ms·
gpt3.5 turbo is (mostly likely) Curie which is (most likely) 6.7b params. So, yeah, makes perfect sense that it can't compete with a 70b model on cost.
by haxton 3y ago
gpt3.5 turbo is (mostly likely) Curie which is (most likely) 6.7b params. So, yeah, makes perfect sense that it can't compete with a 70b model on cost.
- ronyfadel 3y agoIt still does a much better job at translation than llama 2 70b even, at 6.7b params
- two_in_one 3y agoIf it's MOE that may explain why it's faster and better...
- yumraj 3y agoMOE?
- sarthaksrinivas 3y agoMixture of Experts Model - https://en.wikipedia.org/wiki/Mixture_of_experts https://en.wikipedia.org/wiki/Mixture_of_experts
- csjh 3y agoIs there a source on that? I've never seen anyone think it's below even 70B
- why_only_15 3y agogpt3.5 turbo is a new model, not Curie. As others have stated, it probably uses Mixture of Experts which lowers inference cost.
- jiggawatts 3y agoI thought it was fairly well established that GPT 3.5 has something like 130B parameters and that GPT 4 is on the order of 600-1,000
- avion23 3y agoI remember: - gpt-3.5 175b params - gpt-4 1800b params
- JackRumford 3y agoThese sites say 154B: https://www.ankursnewsletter.com/p/gpt-4-gpt-3-and-gpt-35-turbo-a-review https://www.ankursnewsletter.com/p/gpt-4-gpt-3-and-gpt-35-tu... https://blog.wordbot.io/ai-artificial-intelligence/gpt-3-5-turbo-vs-gpt-4-whats-the-difference/ https://blog.wordbot.io/ai-artificial-intelligence/gpt-3-5-t...