Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
redskyluan
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
11 ms
·
91.
▲
by
redskyluan
3y ago
You are right, most LLM queries won't be exact match, so you will need semantic search for hit those similar but not same questions
92.
▲
Yet Another Cache, but for ChatGPT
(zilliz.com)
2 points
by
redskyluan
4y ago
|
1 comments
93.
▲
by
redskyluan
4y ago
We build a cache for cache for ChatGPT, see how it can save your money and time when build application on top of LLMs
94.
▲
by
redskyluan
4y ago
what about the https://huggingface.co/facebook/opt-66b ? I thought the opt series can be used in production
95.
▲
Does ChatGPT Need a Cache?
(github.com)
2 points
by
redskyluan
4y ago
|
1 comments
96.
▲
by
redskyluan
4y ago
do you guys think it is good idea to have a cache layer on top of LLMs?
97.
▲
by
redskyluan
4y ago
what will be the target user of this service?
98.
▲
by
redskyluan
4y ago
I really like this idea. did you plan to add more oss projects into it?
99.
▲
by
redskyluan
4y ago
would you please update the milvus score in your benchmark with the latest milvus, thanks
100.
▲
by
redskyluan
4y ago
did you try to serve billion level vectors, on both projects you mentioned? Then you can define what is overengineered. And what about ScaNN and GPU index? the flexibility to support multi index type is a actually a big plus
101.
▲
by
redskyluan
4y ago
We should probably try to implement a PDF search demo on top of Milvus.. LOL
102.
▲
by
redskyluan
4y ago
See https://milvus.io/
103.
▲
by
redskyluan
4y ago
FAISS is simply a library. You will need a system Like Milvus for huge data amount
104.
▲
by
redskyluan
5y ago
cool idea, what skillset are you guys looking for?