Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
nl
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
9 ms
·
211.
▲
by
nl
3mo ago
> theoretically possible I mean I guess, but not in a performant way if there are ever any hardware failures. And with 100K GPUs there are multiple hardware failures per day.
212.
▲
by
nl
3mo ago
> Go turn on min_p once it's available post July 27th and most of the problems you describe will go away. This seems both arrogantly dismissive ("you are holding it wrong") and incorrect. Either the OP is using Kimi K3 on
213.
▲
by
nl
3mo ago
> set up the environment sloppily The model used a zero-day exploit to escape, and then multiple chained privilege escalations to escape. That indicates the environment both was hardened against all known attacks and had defenses in dep
214.
▲
by
nl
3mo ago
> wildly negligent to not be running it in a physically-airgapped environment Why should it be physically airgapped? Clients won't be doing that.
215.
▲
by
nl
3mo ago
I don't understand this sentiment at all. Is it a claim that "breaking into Hugging Face's production infrastructure" didn't happen? That it's not actually all that severe? That it was done by hand by OpenAI em
216.
▲
by
nl
3mo ago
Deterministic seeds barely work on a single machine, small scale training run. They just don't work at all on a many month long, 100K+ GPU cluster training run.
217.
▲
by
nl
3mo ago
Great, but that seems a different concern to the auditability of a model. You can take the code for Kimi K3 now, take the training framework from Prime and the data from Olmo, spend some money on RL environments and some more money (!) on
218.
▲
by
nl
3mo ago
I know someone who runs AI training. They will have people who don't understand the distinction between visiting Claude.ai and downloading Claude Cowork. They type the words "setup MCP" into Claude.ai and expect it to automat
219.
▲
by
nl
3mo ago
> Getting a SOTA model in a custom harness requires API pricing or risking an account ban, correct? No, only Anthropic has that policy (and I think even that is relaxed for an unknown period if you use the Claude Agent SDK: https:/
220.
▲
by
nl
3mo ago
> because U.S. open weight model makers must follow the frontier labs’ terms of service, they (1) are worse than Chinese alternatives and (2) end up distilling the distillation, just with a detour through Chinese labs. Wouldn’t it be bet
221.
▲
by
nl
3mo ago
Have you ever worked with a non-programmer and helped them setup their AI workflows? You install MCP connectors, specific skills, work around model/harness quirks, set security boundaries etc. It's a lot of work, and most people w
222.
▲
by
nl
3mo ago
I'm all for open models, but people seem to misunderstand what they are. They aren't the same thing as open source code! > open weights, open code and open data Even if you have all these things you still can't replicate a
223.
▲
by
nl
3mo ago
> China's strategy of spending billions on training these models and open sourcing these models away is strategic - they want to kill the US LLM industry at any cost. Why is it when Anthropic and OpenAI spend billions trying to bea
224.
▲
by
nl
3mo ago
That's just not true. You can absolutely have terms of service that are illegal, and the government can enforce them.
225.
▲
by
nl
3mo ago
Sure, but her writing wasn't in any way political.
226.
▲
by
nl
3mo ago
> having PhDs doing night shift lab tech work for pennies I don't know why people keep bringing this up as though it is surprising. In almost any field other than AI PhDs are underpaid on average . There are many, many bio PhDs wor
227.
▲
by
nl
3mo ago
No reason to think it will. Paul Allen died after donating the money to create AllenAI and I don't think there are any links. Hopefully it somehow works out though!
228.
▲
by
nl
3mo ago
> Meta isn't investing in frontier big models anymore Yes they are. Meta Muse is their attempt. It's below frontier performance at the moment but they are spending on getting there.
229.
▲
by
nl
3mo ago
Llama 4 was a bad architecture. Meta Spark is moderately promising but of course closed source.
230.
▲
by
nl
3mo ago
AllenAI is great, but they don't have the budget or remit to build large models.
231.
▲
by
nl
3mo ago
> If the path that was taken to commit this is full of "oops" and "fix" messages great way to encourage people to rebase then!
232.
▲
by
nl
3mo ago
Undecidable isn't uncomputable. "Computable" can mean probabilistic, and classical computers can function over probability distributions just fine.
233.
▲
by
nl
3mo ago
I wouldn't be too fixated on the specific numbers in that post. Anthropic was extremely capacity constrained at that point. They still are but not to that extent. I'd note that OpenAI offers 24 hour caching. I'd be surprised
234.
▲
by
nl
3mo ago
It's not one prompt, but here is a parametric rod connector: Use SCAD and design a connector for square rods. The rods are 18.2 mm square. I want to connect two end-to-end. .. make if the bolt holes are created optional
235.
▲
by
nl
3mo ago
I think that applies to military involvement abroad generally. If you are dropping bombs on someone I'm unconvinced the use of AI will make them like you more or less.
236.
▲
by
nl
3mo ago
I've been doing a lot of 3D design in Codex GPT 5.5 (I found Opus 4.7 wasn't as good - haven't experimented much with 4.8 or Fable). OpenSCAD is a parametric CAD programming language, and the models know it well. The biggest
237.
▲
by
nl
3mo ago
It's been very successful at frontier math tasks - a bunch of the Erdos questions have been solved by it - more than any other model. https://www.erdosproblems.com/
238.
▲
by
nl
3mo ago
Their methodology isn't published. Its widely accepted[1] that it runs the same query through the model in parallel and then has a model that either selects the best answer or synthesizes an answer from the multiple ones generated. I b
239.
▲
by
nl
3mo ago
The source is the GPT 5.5 System Card: > We generally treat GPT-5.5’s safety results as strong proxies for GPT-5.5 Pro, which is the same underlying model using a setting that makes use of parallel test time compute. As noted below, we s
240.
▲
by
nl
3mo ago
I like this format: "I love Lean because <abc>. I found it failed in <xyz> case because <123>. I created a thing <blah> which handles that like this: <ahhh>. I'd love feedback! It's open source h
More ›