Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
stbenjam
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
stbenjam
5d ago
It is a money grab from Amazon I'm sure. They want to make a deal with meta to get paid for access. Maybe also worried about liability I imagine on accidental unauthorized purchases?
2.
▲
by
stbenjam
6d ago
> No special cases introduced. All properties preserved. I don’t actually know if this is LLM-generated, but phrasing like this is weirdly triggering to me now
3.
▲
by
stbenjam
7d ago
Everyone sharing the 513 ways you can already do this is entirely missing the point. Standards matter and there's very few things in the AI world everyone agrees on and this is one of them. .agents/skills is another and anthropic
4.
▲
by
stbenjam
11d ago
The example of the plant watering business is the hellscape future where humans have to deal with agents bugging them to sell them their services constantly. We're already seeing the first hints of this with iLands. And the "hir
5.
▲
by
stbenjam
12d ago
https://github.com/stbenjam/skills/tree/main/plugins/hype I have a plug-in to do this. I don't know if it's effective but Claude said it was genuinely helpful (obviously would say that abo
6.
▲
by
stbenjam
18d ago
By that logic, I think I'll create a new language called Pascal, or maybe Ada...
7.
▲
Show HN: Skillsaw – Linter for Context
(github.com)
1 points
by
stbenjam
18d ago
|
0 comments
8.
▲
Are we measuring AI coding ability wrong?
(bitbin.de)
1 points
by
stbenjam
22d ago
|
0 comments
9.
▲
by
stbenjam
26d ago
The risk here is wildly overstated, prompt injection risk is becoming vanishingly small with the latest frontier models. I would not run an OpenClaw with full access to my bitwarden, but it certainly has some logins available to it, and can
10.
▲
by
stbenjam
26d ago
Ah sorry, for OpenAI it does seem automatic with ID verification. Anthropic's seems more picky.
11.
▲
by
stbenjam
1mo ago
Apply, not sign up. It is not automatic and most will be refused.
12.
▲
by
stbenjam
2mo ago
$80 sounds extremely low for what you're describing - are you on API token plans? I have had some $3,000 token days - even without Fable. I don't see how this is sustainable. My personal 20x plans get so much usage for so cheap.
13.
▲
by
stbenjam
2mo ago
Pi is ok, but I dislike needing plugins to do anything useful. Subagents, MCP, /goal ... none of these things should be plugins.
14.
▲
by
stbenjam
2mo ago
My comment about legal reasons was legal reasons surrounding employee law suits from being laid off. At will employment in most US states allow you to state no reason, that's safer than offering any reason at all. A specific reason f
15.
▲
by
stbenjam
2mo ago
> I hope the big labs will start using this benchmark in their RL pipelines. Labs do not train on benchmark data (allegedly). They can train on similar problems, but benchmarks have specific strings in them that labs are supposed to be
16.
▲
by
stbenjam
2mo ago
I assume it is legal reasons.
17.
▲
by
stbenjam
3mo ago
I know tftp is still in wide use, I wonder if there's things out there looking for stuff that's less common like NNTP, finger servers, etc
18.
▲
by
stbenjam
3mo ago
Oh I didn't hear about this. That sucks. The stock firmware is terrible. CrossPoint makes it usable.
19.
▲
by
stbenjam
3mo ago
I adore my XTEink 4 with the crosspoint firmware. Best small form factor ereader
20.
▲
by
stbenjam
3mo ago
Elon's rhetoric doesn't really match the model's behavior. It is willing to criticize Elon and argues against many of the insane right way points he tries to make.
21.
▲
by
stbenjam
3mo ago
The gulf is bridgeable. The problem is that a lot of people are building agents without strong enough judgment layers around them. Work that can be verified with reasonable accuracy are the sweet spot right now.
22.
▲
by
stbenjam
3mo ago
No need to be defensive. If you ask an epidemiologist, they would almost certainly agree that it is essentially a marker of sexual activity at this point. It is transmissible even with condoms, it has many strains, it is widely prevalent, a
23.
▲
by
stbenjam
3mo ago
I guess you'd rather downvote than answer the question.
24.
▲
by
stbenjam
3mo ago
Chinese models are almost certainly cheating on benchmarks, I would bet if you saw the training data that the benchmark canaries are in there. GLM may be a good model in general but it s benchmaxxed and definitely not as good as Opus 4.8.
25.
▲
by
stbenjam
3mo ago
You realize that the companies listed employ many of the core open source maintainers for large projects? It is project-specific, but 80% of Linux kernel development is from paid corporate employees. Similar for kubernetes. All the load b
26.
▲
by
stbenjam
3mo ago
For some reason not really talked about in mainstream medicine for straight men. It makes no sense. Very safe vaccine and you're eligible into your 40's to get it. Everyone sexually active probably has some strains but not all.
27.
▲
by
stbenjam
3mo ago
I hear a lot of complaints about bun but nothing concrete about what broke in the migration. You are also assuming one prompt, and then arguing against your assumptions with zero evidence. It is lazy arm chair criticism.
28.
▲
by
stbenjam
3mo ago
How can an agent use these tokens then? If it sources the file can't it just read the env? It also sounds like it is missing the important step of keeping the LLM credentials from the agents themselves. For example my GCP creds have a
29.
▲
by
stbenjam
3mo ago
There's a new standard in progress for skills over MCP https://modelcontextprotocol.io/community/working-groups/ski...
30.
▲
by
stbenjam
3mo ago
I like my Nespresso thank you very much.
More ›