Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
gavinboston
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
2 ms
·
1.
▲
by
gavinboston
11d ago
That's one good way to reduce the risk, but I don't think it eliminates it. Even with the same model and the same input, the output is inconsistent. And what I've observed is that as the size of the input and output grows, th
2.
▲
by
gavinboston
11d ago
> I am in the process of attempting to have AI run my business. I'm actually making very good progress, but it's happening in pieces Very cool! What's been the hardest part? Have you successfully automated non-trivial comm
3.
▲
by
gavinboston
11d ago
Well, it's a real example! No, it's not a true story that I'm aware of, but there are plenty of examples of real chatbots run amok. I refactored the tool out of my application and it's available now at https://
4.
▲
by
gavinboston
12d ago
In our new world of non-deterministic output (that's why we love LLMs! they say such helpful/agreeable/sometimes wrong stuff!), I think CI won't be sufficient. CI is in the realm of Quality Control; when I build the thin
5.
▲
by
gavinboston
1mo ago
Do you have a solution for degradation in accuracy when compiling larger amounts of llm-produced text? I am also building LLM knowledge/memory systems and I've been surprised how bad LLMs are, even SOTA models, at summarizing non-
6.
▲
by
gavinboston
3mo ago
Cool project! I haven't seen that OpenRouter workflow yet (sign into OpenRouter and it creates an API key that your app can use), that looks like an interesting pattern to investigate. My company recently built a tool that is closer to