4 ms·
What I mentioned doesn't depend on how LLMs work, the end result is the same (retrieving useful inputs to pass to your LLM). Just meant that a lot of people can
by halflings 3y ago
What I mentioned doesn't depend on how LLMs work, the end result is the same (retrieving useful inputs to pass to your LLM).
Just meant that a lot of people can just do this in-memory or in ad-hoc ways if they're not too latency constrained.