Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
retrovrv
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
retrovrv
1y ago
Your best bet is having an account on AWS Bedrock & Vertex AI so you're able to route your request to the same model (such as claude-sonnet-4) but on a different provider.
2.
▲
by
retrovrv
1y ago
there's portkey that we've been working on: https://github.com/portkey-AI/gateway
3.
▲
by
retrovrv
1y ago
we essentially built the gateway as a service rather than an SDK: https://github.com/portkey-AI/gateway
4.
▲
by
retrovrv
1y ago
just wanted to say, fantastic note. thanks for sharing!
5.
▲
by
retrovrv
2y ago
thank you! we don't have a strong Evals module within our Prompt Studio at the moment. So there's no straightforward way to do that. however, we do have one of the better modules for applying guardrails on live AI traffic and sett
6.
▲
by
retrovrv
2y ago
lol. i get you - i think the 3rd point you shared - it's not exactly about CI issue - CI is already as fast as it can be. It's just that a prompt is a critical part of your AI app and any change to it needs to go through a few hoo
7.
▲
Show HN: Prompt Engineering Studio – Toolkit for deploying AI prompts at scale
2 points
by
retrovrv
2y ago
|
4 comments
8.
▲
by
retrovrv
2y ago
This is actually pretty cool. Would love to try it!
9.
▲
by
retrovrv
2y ago
I'm affiliated with Portkey, so can answer who would need such a proxy/gateway: Sidenote: Arch is def interesting! A typical user we've seen at Portkey is a mid or a large size eng org where a central "Gen AI team"
10.
▲
by
retrovrv
2y ago
came across this guide earlier - valuable insights. thanks for sharing!
11.
▲
by
retrovrv
2y ago
there's an open source ai gateway - https://github.com/Portkey-AI/gateway
12.
▲
by
retrovrv
3y ago
Phenomenal to see how Ragas has progressed. Congratulations on the launch
13.
▲
by
retrovrv
3y ago
https://portkey.ai/ - is also a nice addition to the list
14.
▲
by
retrovrv
3y ago
Thank you! We have built out the cache system -- we do both simple caching (matching the request strings 100%) and also do semantic caching (returning a cache hit for semantically similar requests). More here - https://portkey.ai
15.
▲
by
retrovrv
3y ago
Pretty excited to announce this! While there are some popular and awesome AI gateways out there, like litellm, bricksai - none are written in TS, and for the TS ecosystem. Looking forward to the community's feedback
16.
▲
OpenAI Model Deprecation Guide – GPT3 is shutting down
(portkey.ai)
2 points
by
retrovrv
3y ago
|
0 comments
17.
▲
by
retrovrv
3y ago
I really liked request_format param to enforce JSON outputs, and seed param to enforce deterministic outputs - both I've already started to use.
18.
▲
by
retrovrv
3y ago
Thanks for sharing my blog here! Quick notes on the analysis: - This is based on data from 100+ organizations globally, doing million+ requests a day via Portkey.ai. - I've randomly sampled 10,000 requests for both GPT 3.5 & 4 each
19.
▲
GPT-4 is Getting Faster
(portkey.ai)
3 points
by
retrovrv
3y ago
|
0 comments
20.
▲
Open Source AI Gateway
(github.com)
2 points
by
retrovrv
3y ago
|
0 comments
21.
▲
by
retrovrv
3y ago
Super cool! Looks quite intuitive, especially for function calls.
22.
▲
by
retrovrv
3y ago
There are quite a few LLM monitoring tools in the market. But for monitoring (or evaluating) RAG systems, I found Ragas to be the most helpful: https://blog.langchain.dev/evaluating-rag-pipelines-with-rag...
23.
▲
by
retrovrv
3y ago
There are some programs that match founders, bring in VCs and overall facilitate "Networking" - those are tagged as such. Hope that helps!
24.
▲
by
retrovrv
3y ago
Looks very interesting
25.
▲
by
retrovrv
3y ago
These guys built a few Figma plugins, made money from that, and also got acquired by Figma: https://diagram.com/
26.
▲
by
retrovrv
3y ago
Code interpreter, though not exactly a plugin, has become my default mode of interacting with ChatGPT.
27.
▲
by
retrovrv
3y ago
A lot of thought has been put into this. Quite like the dashboard design, and the simplicity of using this. Congrats on the launch!
28.
▲
Ask HN: Where to Host Llama 2?
3 points
by
retrovrv
3y ago
|
2 comments
29.
▲
by
retrovrv
3y ago
My first question as well, especially considering how fast it is. But understand if you can't reveal. Would still be cool to learn technical details.
30.
▲
by
retrovrv
3y ago
To add to what others have pointed out, if you have a good following on any social platform and can push your product there, putting it behind something like Gumroad is also an option.
More ›