Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
andrewlu0
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
31.
▲
by
andrewlu0
4y ago
Thanks for the feedback Sébastien! 1. Are you adding documents through the API with URLs? URL parsing is only supported through the UI right now, through the API it will just add that url as the metadata. 2. Will add md support soon! 3/
32.
▲
by
andrewlu0
4y ago
Hey! Loving the experience so far - we started the project before pgvector was supported on Supabase, and we want to support hybrid search as well (don't think pgvector supports this?)
33.
▲
by
andrewlu0
4y ago
Yea! Our FE is React/Next and BE is a mix of Supabase, Pinecone, BigQuery and QStash. and mix of HF endpoints and PoplarML in our batch for some embedding models
34.
▲
by
andrewlu0
4y ago
Right now typical usage is < 100 MB per org, but theoretically it could infinitely scale with auto-scaling. And yes heres a new link: https://discord.gg/JCqEzCZ4FM , should be updated on our site as well
35.
▲
by
andrewlu0
4y ago
It should let you trial without a credit card - let me know if thats not working!
36.
▲
by
andrewlu0
4y ago
We might try out both options here - if you check now you'll see we changed it :)
37.
▲
by
andrewlu0
4y ago
We think our current product makes it alot easier to work with documents/embeddings to help build that initial prototype, and once its deployed, tools for versioning and logging. One example use case is ingesting previous support chats
38.
▲
by
andrewlu0
4y ago
Yes, you'll have to bring your own API keys
39.
▲
by
andrewlu0
4y ago
Hi, when you create an endpoint it should give you API examples based on the deployed app. Also docs here: https://docs.baseplate.ai/api-reference/completions . Hope this helps!
40.
▲
by
andrewlu0
4y ago
For standard datasets we use OpenAI ada embeddings, for hybrid its instructor + SPLADE. The HyDE toggle in the context variable feeds the query to a prompt first ("Generate a document that answers..."), before embedding. I think i
41.
▲
by
andrewlu0
4y ago
I feel like this should work? If the original code is too long you may just need a clever chunking strategy
42.
▲
by
andrewlu0
4y ago
Trying to understand your use case more, but if you already have a corpus of example code in JS and in Python, why the need to use the model to do the transformation?
43.
▲
by
andrewlu0
4y ago
Right now we have APIs to directly query the data, and you can export logs as csv. If there's any integration or export feature you'd like let us know!
44.
▲
by
andrewlu0
4y ago
We added that just to prevent spam, but we do have the 7 day trial. The tour is pretty quick and you can cancel if you don't see enough value!
45.
▲
by
andrewlu0
4y ago
Neat! just fyi i think that link is broken
46.
▲
by
andrewlu0
4y ago
Yes, we support both those models!
47.
▲
by
andrewlu0
4y ago
Layup and Tennr are both using our team plan. We cited these two since they gave us quotes to use on our website. They are great products that we are excited to support! We also have several teams outside of YC on the team and pro plans.
48.
▲
by
andrewlu0
4y ago
Our original idea was based on an internal tool at Google from this paper https://arxiv.org/abs/2203.06566 , but we took it in a different direction after some early user interviews. Not going to pretend we were using G
49.
▲
by
andrewlu0
4y ago
Something we think about a lot! Under the hood, we've built an orchestration layer that keeps a database, a vector database, and storage in-sync. For the startups we work with, this has to scale to thousands of documents with high thro
50.
▲
by
andrewlu0
4y ago
Yes, and actually you don't need to do steps 2 and 4/5 yourself, it can be built on Baseplate, and you just need one API call to our endpoint to get the result. You can send us the blog post through our API as a file or copy/
51.
▲
by
andrewlu0
4y ago
No reason really it was just a random kinda large PDF I found online - we're not targeting a specific market and think we could provide value across many different industries
52.
▲
by
andrewlu0
4y ago
Yep! Should be supported - we have standard CRUD apis for managing data/documents
53.
▲
by
andrewlu0
4y ago
This is something we'll need to figure out since we don't have specifics on how the image is fed into the model. An example (very basic) use case we thought of is maybe a database of apartment listings & photos, and being able
54.
▲
by
andrewlu0
4y ago
we definitely took some inspiration from linear :)
55.
▲
by
andrewlu0
4y ago
Thats actually a pretty good idea, we'll consider it! One thing is since our hybrid datasets are based off instructor-large embeddings, we'll later offer the ability to set the embedding "instruction", which can be tweak
56.
▲
by
andrewlu0
4y ago
Haha, the support one is great :)
57.
▲
by
andrewlu0
4y ago
Yes, currently for the context data it needs to be stored on Baseplate. But we are exploring that direction of tools (a logical next step), where it could do a Google search or custom API call etc..
58.
▲
Launch HN: Baseplate (YC W23) – Back end-as-a-service for LLM apps
177 points
by
andrewlu0
4y ago
|
104 comments