4 ms·
Latency is actually very low as-is, with GPT-4-Turbo! But other than that, yes, we do have plans - using real-time RAG. I carried out some research with incred
by swiftlyTyped 3y ago
Latency is actually very low as-is, with GPT-4-Turbo!
But other than that, yes, we do have plans - using real-time RAG.
I carried out some research with incredibly promising results
https://ai88.substack.com/p/rag-vs-context-window-in-gpt4-accuracy-cost https://ai88.substack.com/p/rag-vs-context-window-in-gpt4-ac...