5 ms·
I think both things can be true: new models benchmark higher and eat more tokens.
by cmiles74 3mo ago
I think both things can be true: new models benchmark higher and eat more tokens.
- jbvlkt 3mo agoFrom my experience new models are slower and use more tokens even on questions which gpt 4 answered correctly. It is mostly because newer models tend to be more verbose (even with prompt requesting short answers).
- bcjdjsndon 3mo agoUnless somebody improved on the underlying transformer architecture... Surely AI is smart enough to do it by now