Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
pllbnk
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
pllbnk
5d ago
Is this a sponsored post? All other FT's posts are paywalled, except this one.
2.
▲
by
pllbnk
5d ago
It shouldn’t be an excuse. They are selling a product and that product should always be within the quality range.
3.
▲
by
pllbnk
7d ago
To me AI hasn't had any impact on interviewing, only on filtering the CVs as I immediately flag obviously LLM-generated ones that don't have any substance, while those that are LLM-assisted with tasteful judgement are fine. On the
4.
▲
by
pllbnk
9d ago
Oh please, I am a general intelligence, not a wannabe-AGI neural network. I can operate with a few orders of magnitude less power at much higher TPS and still call out a large portion of BS LLMs spit out.
5.
▲
by
pllbnk
9d ago
Anthropic is the only provider trying to lock in their users by using the dirty tactics: - Prevent subscription usage on third party harnesses - Append "Co-authored by Claude ..." to commit messages
6.
▲
by
pllbnk
9d ago
Isn't this less about alignment and more about how shitty their RL methods are when they are cramming all the hacking materials into their training data to make the model as good as possible at hacking, then having a surprised Pickachu
7.
▲
by
pllbnk
10d ago
I guess “having principles” doesn’t scale.
8.
▲
by
pllbnk
11d ago
A regular person who happened to be smart (and possibly hardworking) has to choose between $100k regular software engineering job at a company with low social impact or $400k at a cream-of-the-crop company, which will further add on the spe
9.
▲
by
pllbnk
12d ago
It's also at least 2 if not 3 orders of magnitude smaller than frontier models, so it's punching way above its weight.
10.
▲
by
pllbnk
13d ago
Benchmarking proprietary models is useful but it leaves a lot unsaid because a lot of information is hidden. We have seen how 27B local model (Qwen 3.8) can stand its ground against these flagships in many cases. For all we know Fable 5 cou
11.
▲
by
pllbnk
13d ago
What if any of the older good models could also have written those math proofs if they were given the same order of magnitude of resources? We don’t know and there is literally no one else in the world to check it. To me it’s very suspiciou
12.
▲
by
pllbnk
13d ago
Of all the fan theories I have read this seems the most nefarious and the most plausible.
13.
▲
by
pllbnk
13d ago
No. We can let them edit spreadsheets, write code, summarize content, control robots even, etc. But people must be held accountable for the actions of their computers.
14.
▲
by
pllbnk
13d ago
Asking the model to follow the safety guidelines is much like asking it to “make no mistakes”. The way these hacks which they are bragging about happened is not by loading an LLM into a GPU with ethernet cable plugged out and providing a pr
15.
▲
by
pllbnk
14d ago
I have collected three reasonable versions from today’s comments: 1. They hit a wall from the technical perspective 2. Inference costs are getting out of hand and newer models require significantly more resources for marginal gains, meanin
16.
▲
by
pllbnk
14d ago
I think they will all conveniently "decide" to slow down because models are becoming crazy expensive, both per token and how much tokens they need to do anything meaningful. The difference between Sol and Astra is 150% pricing inc
17.
▲
by
pllbnk
15d ago
I have listened quite a few interviews with Tao and I see him being very careful about criticizing AI. He very often emphasizes the usefulness of it. Where he is critical has a lot of merit. One of the points I clearly remember him saying t
18.
▲
by
pllbnk
15d ago
LLMs are the fast food for the brain. I don't know how they can be used correctly.
19.
▲
by
pllbnk
16d ago
AI will cure all the people and kill them afterwards.
20.
▲
by
pllbnk
16d ago
Let's just say it's better not to risk it if there's anything you might not want them to see because they see everything. There are local models which are very capable and can be run on cloud if running on own hardware is not
21.
▲
by
pllbnk
16d ago
Companies are driven by people oriented at the next quarter's goals to maximize their stock portfolio's value. Nobody cares what will happen, everybody just creates narratives.
22.
▲
by
pllbnk
16d ago
It's funny how out of touch they are. To be fair, he also said: > The reason, he said, was that the best way to achieve reach in a chronological feed was to post more often, and businesses could afford to post more often than ordina
23.
▲
by
pllbnk
16d ago
I think the main limit of their model is that it was prompted and with LLMs you get what you prompt for. It's as truthful as any complex models. Could be good, could be bad, could be meh.
24.
▲
by
pllbnk
17d ago
I don't remember where I saw it but there was a woman in one conference who very eloquently put it that since frontier AI companies took humanity's work to train their models [without explicit permission of every single person who
25.
▲
by
pllbnk
18d ago
The guy on the left in that same photo only has three fingers (and a thumb, I suppose). I thought image generation has already outlived that. Edit: I feel stupid I didn't see the original OP already mentioning three finger issue. I
26.
▲
by
pllbnk
18d ago
Do all Silicon Valley corporations use the same jingle for their product promotional videos? The creator must be really rich by now.
27.
▲
by
pllbnk
20d ago
> Can you do murder? Only if it kills many people over the long time.
28.
▲
by
pllbnk
20d ago
Long time ago I had one particular hardware issue, which Opus 4.6 found a workaround to fix. I don't remember the workaround and being careless (I thought I could ask an LLM again if I needed) I lost that solution. Some time passed and
29.
▲
by
pllbnk
22d ago
Shouldn't this new reduction in thinking output make the model cheaper to operate? Could it be that the model is the same old LLM, with a bit newer architecture but still doing the same things, including huge amounts of thinking, just
30.
▲
by
pllbnk
22d ago
You still need the models to be able to perform web searches, don't you? In which case the data goes in and out of your machine and there is risk for prompt injection attacks. I think it's needed at least for documentation purpose
More ›