Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
gr_norm
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
61.
▲
by
gr_norm
2mo ago
> Distillation does not allow the CCP to obtain equivalent or superior AI capabilities to the US, but it can bring the Chinese frontier to within a few months of the US frontier A message to their investors, it would seem. "They cau
62.
▲
by
gr_norm
2mo ago
I don't really understand how they can argue the security angle with a straight face. It's not like GLM 5.2 is a slouch. I've seen it do things like exploit an IDOR issue when I was experimenting with a quick-and-dirty web au
63.
▲
by
gr_norm
2mo ago
> Open-weights models that don’t have dangerous capabilities are a public good Note the hedging against 'dangerous capabilities'. Undoubtedly, all the useful ones trigger this condition in Anthropic's eyes. The rest of the
64.
▲
by
gr_norm
2mo ago
> I’m not even going to get into how you could provably transform brute force propositional logic into efficient algorithms. > The entire problem is showing that an efficient/reliable program actually implements those rules. Reli
65.
▲
by
gr_norm
2mo ago
It is true that finding the correct specification is a formidable task; knowing what correctness even means is arguably most of the difficulty of programming. However! "Moving bugs up from programs to types" isn't how this sh
66.
▲
by
gr_norm
2mo ago
To maximize personal influence and wealth, of course. These days, many with power don't seem to care much about uplifting society so long as they get theirs. Hopefully Anthropic eats crow here, lest their wish is fulfilled that we all
67.
▲
by
gr_norm
2mo ago
Many who rush to the defense of AI companies' marketing departments seem to take criticism personally, as if not buying into it all hook, line, and sinker is an affront to them. A foreseeable consequence of becoming cognitively depende
68.
▲
by
gr_norm
3mo ago
Login-walled for me. Is there a Reddit proxy like Nitter?
69.
▲
by
gr_norm
3mo ago
I could fully see them thinking the incident disclosed yesterday would have made them look good ("wow, OpenAI's models are so capable!"). That it didn't occur to them to discuss specific preventative measures to be taken
70.
▲
by
gr_norm
3mo ago
Yeah, open weights hasn't fully caught up yet, but it's getting very close. And it's certainly passed the point of being reasonably interchangeable with the frontier for (programming) work. Add in the benefits of not being ru
71.
▲
by
gr_norm
3mo ago
The idea that the Chinese labs cannot make progress except by copying superior American products is just prejudice against the former and exceptionalism of the latter at play. Even the OpenAI top brass have admitted otherwise [1]. China is
72.
▲
by
gr_norm
3mo ago
HF need not be party to it at all, beyond being the victim. I suspect the hack is real; I have observed GLM 5.2 being able to discover similar vulnerabilities in web applications I'm hosting (which I've then fixed!). At the same t
73.
▲
by
gr_norm
3mo ago
The timing after the release of GLM 5.2 and Kimi K3 is quite convenient, too, as an angle for regulatory quashing of open-weights models just as they're entering the mainstream conversation around usurping the American frontier labs. I
74.
▲
by
gr_norm
3mo ago
They've been saying so from the beginning, and yet did not take the basic precaution of airgapping their off-the-leash model while it's been instructed to succeed at a hacking benchmark by any means necessary. So which is it? I _w
75.
▲
by
gr_norm
3mo ago
Yeah, agree on all counts. I'd give them leeway if they were still scrappy startups, but they have entire countries' worth of resources at their disposal and the best of the best on their payroll. No excuses at this point for oops
76.
▲
by
gr_norm
3mo ago
Why is a machine running these sorts of hacking benchmarks not airgapped? That seems a basic precaution, if OpenAI believes what they're selling. I mean, stuff like this is done for CTFs played by humans, too, to rule out collateral da
77.
▲
by
gr_norm
3mo ago
This is already detestable given how much they're charging, but how long before LLM ads become difficult to distinguish from the main output? Remember how Google ads started out? And for all their faults, Google _after_ 20 years of max
78.
▲
by
gr_norm
3mo ago
Yeah. It's something I can do myself in a couple seconds if I want, also on more varied SVG scenes. If this is going to be a benchmark people turn to I'd like to see more effort put into it than just a one-sentence prompt.
79.
▲
by
gr_norm
3mo ago
Yup, fair's fair. Anything else stinks of 'rules for thee but not for me' (a maxim the frontier labs seem worryingly happy to apply, on several counts).
80.
▲
by
gr_norm
4mo ago
Many words expended here to avoid asking an answerable question. Which complaints? Must I recapitulate the last 70 years of history of politics around science in the United States? You have fixated on a singular line from the article rather
81.
▲
by
gr_norm
4mo ago
Sorry, I'm not in the business of responding thoughtfully to low effort questions slung rapid-fire over the fence, which moreover take nothing I've said into account. Case in point, this comment; it seems innocuous but would take
82.
▲
by
gr_norm
4mo ago
That is an enormous budget cut; it is exactly an evisceration. And the general trend is to cut both science and its application. Trials for life-saving treatments _by biotech companies attempting to commercialize them_ are being halted or c
83.
▲
by
gr_norm
4mo ago
You and people you know will lead worse and less fulfilling lives due to this. You, or someone you love, will likely die of causes that would have been preventable without this destruction of domestic science. Academic culture undoubtedly n
84.
▲
by
gr_norm
4mo ago
I wouldn't normally reply again, but please try and take my arguments in the good faith and appeal to the heart and humanity that they were given. I think you're quite capable of reading what I've written and responding to it
85.
▲
by
gr_norm
4mo ago
Anyone with kids could tell you that this is an ethically vacant position to hold. Particularly given the social network effects of not being on these platforms and the addiction engineering that goes into keeping you on them, especially at
86.
▲
by
gr_norm
4mo ago
This is tired and reductivist reasoning deployed only by those rationalizing wrongs, and immediately recognizable by nearly everybody as such. My kids could tell you what's wrong with this thinking. I believe that working a job where I
87.
▲
by
gr_norm
4mo ago
Nope, sorry, I have absolutely had the option to do something similar and emphatically declined. I generally don't care to tell anyone this, either, outside the rare instances when it organically comes up as it did here. I want to see
88.
▲
by
gr_norm
4mo ago
I am completely willing to forgive Meta (and Palantir etc) employees who quit their job and donate their blood money (all wages above some low multiplier of median US SWE salary, adjusted for cost of living) to a reputable charity of their
89.
▲
by
gr_norm
4mo ago
You need tools sufficient to do the job in an economical way, optimizing for both cost and quality. That is what 'best' means. We don't give every engineer all the resources under the sun, only what is appropriate. I suspect
90.
▲
by
gr_norm
4mo ago
> Since it is the unverified SMP config of the kernel I don't disagree with your point (formal verification does not rid you of all bugs), but this is not the subject of the linked issue. This was a bug in an unverified path.
More ›