3 ms·
They are already solving the problem with search engines, they're just using an LLM as a first pass to create better embeddings to run a similarity match on fir
by quixoticaxolotl 2mo ago
They are already solving the problem with search engines, they're just using an LLM as a first pass to create better embeddings to run a similarity match on first. The difference in latency is likely made up for in accuracy.