3 ms·
Also, just talking about the raw input / output raw cost leaves out half of the equation. You also have to think about how a given providers caching strategy be
by elliottshort 2mo ago
Also, just talking about the raw input / output raw cost leaves out half of the equation. You also have to think about how a given providers caching strategy behaves, and how many tokens a given model actually uses to complete a task on average.
Just because a model has a higher input / output token cost doesn't necessarily always mean it's going to be more expensive to use.