3 ms·
We've done a bunch of work to strip down the context and minimise the output tokens (which tends to be 100x as slow as input tokens). GPT-4o is pretty fast too
by ppsreejith 2y ago
We've done a bunch of work to strip down the context and minimise the output tokens (which tends to be 100x as slow as input tokens). GPT-4o is pretty fast too :)
- penthi 2y agoThanks for the explanation. Can't wait to see the code when you open it up!