Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
preommr
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
preommr
5d ago
Well I think there's an expectation that it's similar to the probability score, something that's outputted by the model itself, and so there's some level of "intelligence" (e.g. being able to recognize if the s
2.
▲
by
preommr
5d ago
> its confidence scores important to note that the "confidence" score is... maybe not what people think it is - kind of useless, and just a convenience step from the probabilities. from the docs: "confidence is a statistic
3.
▲
by
preommr
6d ago
And yet there's none of this uncertainty or caution with his posts that then go on to have massive ripple effects because of his position at Anthropic and the marketting related to claude. I personally have a lot of anger and frustrati
4.
▲
by
preommr
11d ago
This will be insane for tool usage, and probably where the major economics for day-to-day usage will be. The goal is going to be to use llms to distill operations down to some dsl, and pass it into something like Jev.
5.
▲
by
preommr
13d ago
> and you will face criminal and financial penalties for damages caused. I don't understand why this isn't talked about more. We don't need a slowdown. Just double down on prosecuting crimes. Let the companies take the ris
6.
▲
by
preommr
14d ago
What do these discussions even matter when the words don't matter? How many times has AGI been declared already? A bunch of people (e.g. Jensen Huang) have called Astra AGI, for example. The goal posts get moved, everybody's hustl
7.
▲
by
preommr
17d ago
It's crazy how deep Microsoft's bench is (Lean was started there, vscode is another), for everything not directly related to the the ai models (hell, even github for data). So interesting how everything played out, I remember in t
8.
▲
by
preommr
19d ago
These people need to take a vacation and come back in a few months when the vision models get better/cheaper. Astra is already good at taking screenshots and acting on it (part of the agi claims).
9.
▲
by
preommr
22d ago
close enough. The second to last line is "book it" for some tennis thing, and the scene before that has the guy eating the food the ai ordered.
10.
▲
by
preommr
27d ago
> The fact that it is almost an anagram of monarchy is probably a plus for DHH. I spend way too much time online; but it's good to know I am not this terminally online.
11.
▲
by
preommr
28d ago
Documentation was difficult to keep in sync with code because the tools didn't operate at a natural language level. Things like refactoring names, or terminology, or concept changes (foo is now fooGroup that contains bar) were tedious.
12.
▲
by
preommr
28d ago
This conversation again? They go nowhere because people are using wildly different definitions and contexts. There's one already in here about how ai is better than humans at coding. - Yes llms are better at the mechanics of coding - n
13.
▲
by
preommr
29d ago
Disagree, partially. Buy-into storing credentials into environment variables. Then, and this is important - MANAGE YOUR ENVIRONMENTS. You shouldn't have prod level s3, or aws creds accessible openly in your environment. If someone can
14.
▲
by
preommr
1mo ago
> Do you think this is where it stops? This is where it begins. No, this is pretty much where it stops. The models are good enough for the average coding task, and the slop they produce often is in the category of what a bad or careless
15.
▲
by
preommr
1mo ago
wtf is this garbage? I don't think the average C dev isn't aware that the language has pitfalls, especially around memory management. Also a lot of these problems are because C is used everywhere and has been for so long. Not beca
16.
▲
by
preommr
1mo ago
I've had codex delete useful (albeit not directly relevant or perhaps messy wip notes) comments, even though I explicitly have it in my agents.md not to delete comments, and ask for permission if it thinks it should. It deleted the com
17.
▲
by
preommr
1mo ago
So the vagueposting by googlers about Ox Alpha was just... what exactly? Like I get that they have to be careful about comms, but surely senior members of the team can clarify when something is NOT them, when everyone is gosspiing it is the
18.
▲
by
preommr
1mo ago
People have taken 'everything is political' to mean 'politics must be discussed for everything'. I am old enough that I've been burned by both sides of the political aisle, often using the same rhetoric (think of th
19.
▲
by
preommr
1mo ago
> Every software engineer in my company uses Claude code heavily. I find it funny how OpenAI got caught lacking for a very brief window, but it turned out to be a very critical turning point. Like a guy that that's at the top of the
20.
▲
by
preommr
1mo ago
Agents.md are (and probably will continue to be) an ugly band-aid. - new model comes out and a bunch of it becomes obsolete - they get flat out ignored, esp. with larger context windows. The ai just responsds with, "your'e right I
21.
▲
by
preommr
1mo ago
Because it's a separate marketing term. Instead of the CEO mandating that the API server has to be agent compatible (where who knows what that means), they can just say "our product has an MCP". On a technical level, who know
22.
▲
by
preommr
1mo ago
omfg, that's a long post. How is this even possible? I know we're all using AI, but Bun seems like the one singular project where there's just been a crazy increase in the amount of output, a 10x on the 10x. How are they doin
23.
▲
by
preommr
1mo ago
> Why would I want everything reimplemented in this massive binary? Why would Bun be anymore in touch with nuances of all these different technologies then individual projects dedicated to their own speciality? It's funny because th
24.
▲
by
preommr
1mo ago
> GPT-OSS 120B which is nigh useless nowadays: I still think that was a really great model that got overlooked. It was really great in terms of latency/throughput while still being fairly intelligent. I was planning on using it for
25.
▲
by
preommr
2mo ago
People are missing the point if they think this is useless because frontier models keep changing every few months. We really, really need better secondary models that can do things fast and do them cheaply for lots of dumb tasks. Not only b
26.
▲
by
preommr
2mo ago
Of course it's Not Just Bikes. I used to be frustrated at their poorly thought-out arguments until I realized none of it matters. The videos are made for a certain demographic to jerk off about their special interests. The real world i
27.
▲
by
preommr
2mo ago
No it isn't. Your local .env should NOT be shared, but should also be assumed to be leaked at any given time. Security should be done using a secrets manager through the cloud platform that's being used, e.g. AWS'secrets mana
28.
▲
by
preommr
2mo ago
> encountered Epstein-class people doing what they do with the children of their own staff kind of glossing over something pretty major
29.
▲
by
preommr
2mo ago
I pay $20 for codex, use it daily for coding, and still haven't dipped below 50% for weekly usage. I wouldn't even be able to afford buying 29 gigs of ram, or a new video card with hardware prices the way they are now. Maybe if it
30.
▲
by
preommr
2mo ago
I thought the chinese models were cheaper per token, but about the same or more expensive on tasks because they used more tokens for reasoning. Cutting even further, seems like a really big leap.
More ›