Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
santiago-pl
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
santiago-pl
24d ago
rtk gain mechanism is oversimplified. 1 token != 4 bytes for the standard prompt / context window at coding agent.
2.
▲
by
santiago-pl
27d ago
Right now we have those OpenClaw commercial alternatives: Grok Bot, Muse, Viktor. Anything more?
3.
▲
"I'm a father of three who studies the impact of artificial intelligence"
(theguardian.com)
2 points
by
santiago-pl
29d ago
|
1 comments
4.
▲
Poland offers businesses seats on presidential plane for foreign visits
(notesfrompoland.com)
2 points
by
santiago-pl
1mo ago
|
1 comments
5.
▲
by
santiago-pl
1mo ago
I like the idea! Two things I would like to see there: 1. In Poland we don't have Euro - I would love to see a switch / dropdown to change the currency for displayed prices. a) Also I think it will be convenient to for US base
6.
▲
by
santiago-pl
1mo ago
It's simple and it's working! I can't find a GitHub link. I assume it's not open-source. Am I correct? WDYT about building a browser extension based on this?
7.
▲
by
santiago-pl
1mo ago
Did anyone compare it with gpt-6?
8.
▲
Nvidia Agrees to Buy AI Platform Hugging Face for $13B
(wsj.com)
1 points
by
santiago-pl
1mo ago
|
0 comments
9.
▲
Why Cheap Link Building Is the Most Expensive Marketing You Can Buy
(entrepreneur.com)
1 points
by
santiago-pl
1mo ago
|
1 comments
10.
▲
by
santiago-pl
1mo ago
GoModel author here. If you are looking for a self-reproducible benchmark, I run this one periodically. I keep it as honest as possible to know how GoModel stands in comparison to its competitors. Any feedback appreciated via GitHub issues:
11.
▲
by
santiago-pl
1mo ago
GoModel author here. Someone asked me about a benchmark for GoModel during my "Show HN" launch here. A few days ago, I moved it to a separate repository and gave it another run. The benchmark is self-reproducible, so you can run i
12.
▲
Want people to use cheaper AI models? Make expensive models slower
(enterpilot.io)
1 points
by
santiago-pl
2mo ago
|
1 comments
13.
▲
by
santiago-pl
2mo ago
I came up with this idea of artificially slowing down expensive models within a company to incentivize people to use cheaper ones. I'm curious - has anyone tried this at scale?
14.
▲
by
santiago-pl
2mo ago
GoModel author here. I prepared self-reproducible benchmarks and published them on my blog. Might be helpful. I don't think latency is the best metric for comparing AI gateways, as most of the latency comes from the providers. Memory u
15.
▲
by
santiago-pl
3mo ago
If you experience any issues with LiteLLM, you may try GoModel - the AI Gateway I'm working on. It consumes ~60x less resources and is more reliable :)
16.
▲
by
santiago-pl
3mo ago
Actually, LiteLLM is the most popular, it was first product of this type, but there is plenty of more efficient and reliable alternate AI Gateways right now. I've written one - GoModel. https://gomodel.enterpilot.io/
17.
▲
by
santiago-pl
3mo ago
I'm working on the last AI gateway you'll ever try - GoModel :) I recently created a benchmark and it looks like GoModel is the fastest and most lightweight open-source (self-hosted) AI gateway on the market. I tried to make it as
18.
▲
by
santiago-pl
3mo ago
I’ve created an honest, reproducible benchmark of self-hosted AI gateways literally two days ago. I posted it on HN here: https://news.ycombinator.com/item?id=48688213 To be totally transparent, I’m the author of GoModel, b
19.
▲
by
santiago-pl
3mo ago
I mentally treat it as part of the documentation right now.
20.
▲
by
santiago-pl
3mo ago
Currently, I want to keep them close to the documentation, where I link to them directly. The fewer repositories I have, the easier it is to maintain them.
21.
▲
by
santiago-pl
3mo ago
Thank you! Feel free to reach out to me over Discord in the case of any feedback or feature request.
22.
▲
Benchmarking AI Gateways: GoModel vs. LiteLLM vs. Portkey vs. Bifrost
(enterpilot.io)
4 points
by
santiago-pl
3mo ago
|
9 comments
23.
▲
by
santiago-pl
3mo ago
GoModel author is here. This is my attempt to honestly compare GoModel vs LiteLLM vs Portkey vs Bifrost AI gateways.
24.
▲
by
santiago-pl
5mo ago
Giorgi is the semantic caching master at GoModel right now. Let me ping him, and he'll get back to you here.
25.
▲
by
santiago-pl
6mo ago
I'll definitely take a look at GAI myself! I like this beaver(?) at README.
26.
▲
by
santiago-pl
6mo ago
I'm wroking on it full-time right now. It might be challenging, especially when it comes to interactions with video, audio, and image models. I'm just trying to stay on top of what's happening and add new things day by day. A
27.
▲
by
santiago-pl
6mo ago
My thoughts about this: Benchmarking AI gateways properly is harder than it looks. Feature sets differ meaningfully - exact vs semantic caching, cluster mode, guardrails, audit logging - and each carries its own latency cost. What actually
28.
▲
by
santiago-pl
6mo ago
I've released a new version of GoModel (0.1.20) with explicit support for vllm. You can now use it even with a few vLLM instances. Like this: docker run --rm -p 8080:8080 \ -e VLLM_BASE_URL=http://host.docker.internal
29.
▲
by
santiago-pl
6mo ago
That's great news! The AI model ecosystem is changing so fast.
30.
▲
by
santiago-pl
6mo ago
TBH I decided to write GoModel because I needed something like this for my startup, enterpilot, and LiteLLM didn’t meet my needs.
More ›