Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
kingstnap
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
kingstnap
5d ago
There are three big things in this announcement it seems. > In addition to resolving the Navier–Stokes Millennium Prize problem , this model has now resolved more than 100 long-standing open problems across most areas of mathematics. Rum
2.
▲
by
kingstnap
5d ago
I mean the Pi currently makes little to no sense to buy anyway. So locking down parts barely moves the needle. The value proposition is unbelievably bad at spending upwards of $423 CAD for a 16GB Raspberry Pi. Which is literally just the ba
3.
▲
by
kingstnap
6d ago
The headline makes it seem more monetary than it actually is. https://www.read.gov.sg/join-the-readsg-challenge It seems like turning points into money is just a happens to exist sort of mechanic. Enough to write a headline
4.
▲
by
kingstnap
6d ago
According to OpenAI the cut off date for user data was too early for that (one sided evidence, so I'll give this partial consideration). The NYU professor was solving a different problem (no viscosity, aka the Euler equations). This is
5.
▲
by
kingstnap
8d ago
As opposed to taking advice from the glucose wetware?
6.
▲
by
kingstnap
9d ago
Just note that Terry Tao is on the letter. This is a guest post. https://mathandai.org/
7.
▲
by
kingstnap
9d ago
They do, when Luna got 5x cheaper it was directly attributed to some unknown % inference optimization. US labs are quite cut throat about dealing with stuff costing them money (inference). This sort of engineering excellence doesn't al
8.
▲
by
kingstnap
9d ago
The first implies generalization. It's not a test of generalization. It's actually close to the second. "Here are 200 software questions, will drill you on *other stuff* until you can pass exactly these 200. If the other stuf
9.
▲
by
kingstnap
9d ago
It's implicitly trained against. There is like information leakage with researchers messing with the training parameters and checkpoints used. It's not the direct feedback loop of RL but its not far.
10.
▲
by
kingstnap
11d ago
As someone who used to use it a ton, Reddit is such a shadow of its former self that it's insane. It's actually like 90% bots. Their user growth numbers are all fake. The only real people on there are marketers making fake posts a
11.
▲
by
kingstnap
15d ago
Performance problems in theorem provers is an old topic. I remember watching this and it was fun. https://youtu.be/m-iGCCuHBvY [Talk] 10 years of superlinear slowness in Coq (2022)
12.
▲
by
kingstnap
15d ago
This is pretty interesting in a lot of non-surface-level ways. I can see OpenAI pushing for this as a sort of more durable moat compared to the now huge number of agentic harnesses that run on your own machine. This might be getting the foo
13.
▲
by
kingstnap
18d ago
Replace hire with marry and this is *exactly* what politics was globally for all of human history. This is not something "behind us" btw.
14.
▲
by
kingstnap
19d ago
> hey Claude vibeslop me a CUDA Anubis solver" route is on its way to being fundamentally dead. Lmao yeah no. I don't think a little argon2 is going to change shit all. I mean the thesis of Anubis itself is "scrappers are
15.
▲
by
kingstnap
21d ago
In real life, there is always this feedback edge from the results to the methodology. Theoretically it's unscientific to do tweaks like this but in reality this is what actual science is because you need to see the results understand t
16.
▲
by
kingstnap
21d ago
Its also available finally to Pro users! Just took 24 hours.
17.
▲
by
kingstnap
22d ago
The AI labs have out considerable effort in trying to find and patch lean exploits. They explicitly set agents and have them try to prove false. > Daniel used OpenAI internal models to discover new soundness issues in the official Lean k
18.
▲
by
kingstnap
23d ago
Are you guys just posting what you think might end being the link so you can farm upvotes or something?
19.
▲
by
kingstnap
23d ago
This seems to be the announcement w/ the trust me bro numbers. https://x.com/Alibaba_Qwen/status/2094968708288680276
20.
▲
by
kingstnap
23d ago
They are simply measuring in base 1.7 :^) Uber has grown 2 orders of magnitude!
21.
▲
by
kingstnap
24d ago
Hosting is probably ~0% of the cost lol. The costs of the writers is what makes journalism expensive.
22.
▲
by
kingstnap
24d ago
The blog post is gone but I can currently use it in the gemini chat website.
23.
▲
by
kingstnap
25d ago
Do API prices not affect usage limits for subscriptions? They do in Codex.
24.
▲
by
kingstnap
25d ago
Haiku would have to be a banger, with a significant price drop, to make any sense. It's currently priced 33% above Gemini 3.7 Flash, and several multiples of 5.6 Luna.
25.
▲
by
kingstnap
26d ago
I think the jobs thing will be a slowly then all at once thing. Firstly, AI is sort of like this to begin with. It's useless for something until it gets hill climbed and is useful for something. AI agents for chip design was an idiotic
26.
▲
by
kingstnap
28d ago
Man that claude writing is bad. Really bad.
27.
▲
by
kingstnap
29d ago
Smaller models + more effort has strong diminishing returns, especially if your goal is to save money. Sol already lacks judgement. It will absolutely add idiotic tests and comments. Luna is that but worse so if you account for things like
28.
▲
by
kingstnap
1mo ago
Italy too. It's bizarre trying to hop on the trenitalia trains while a bunch of random people are standing on the platform right in front of the doors trying to get a few last drags in before the train leaves.
29.
▲
by
kingstnap
1mo ago
A random blight in Canada in that Bell still hasn't enabled IPv6 for residential plans. You can get 8 Gigabit symmetric home internet from Bell but not IPv6.
30.
▲
by
kingstnap
1mo ago
Whats stopping you? You could buy puts right now. Get a 210 strike put contract and if your thesis is that nvidias current 10 day slide continues you could make some money.
More ›