4 ms·
We have a cloud offering at https://trypromptly.com https://trypromptly.com. We do offer enterprises the ability to host their own vector database to maintain c
by ajhai 3y ago
We have a cloud offering at https://trypromptly.com https://trypromptly.com. We do offer enterprises the ability to host their own vector database to maintain control of their data. We also support interacting with open source LLMs from the platform. Enterprises can bring up https://github.com/go-skynet/LocalAI https://github.com/go-skynet/LocalAI, run Llama or others and connect to them from their Promptly LLM apps.
We also provide support and some premium processors for enterprise on-prem deployments.
- mveertu 3y agoEnterprises can bring up https://github.com/go-skynet/LocalAI https://github.com/go-skynet/LocalAI, run Llama or others and connect to them from their Promptly LLM apps - So spin up GPU instances and host whatever model in their VPC and it connects to your SaaS stack? What are they paying you for in this scenario?
- rsiqueira 3y agoBut, in order to generate the vectors, I understand that it's necessary to use the OpenAI's Embeddings API, which would grant OpenAI access to all client data at the time of vector creation. Is this understanding correct? Or is there a solution for creating high-quality (semantic) embeddings, similar to OpenAI's, but in a private cloud/on premises environment?
- ivalm 3y agoSentence-Bert is at least as good as OpenAI embeddings. But I think more importantly Azure OpenAI model api is already soc2 and hipaa compliant.
- ajhai 3y agoEnterprises with Azure contracts are using embeddings endpoint from Azure's OpenAI offering. It is possible to use llama or bert models to generate embeddings using LocalAI (https://localai.io/features/embeddings/ https://localai.io/features/embeddings/). This is something we are hoping to enable in LLMStack soon.