Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
Philpax
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
Philpax
10d ago
> It's not going to happen accidentally. https://transformer-circuits.pub/2026/emotions/index.html Whether these are like "our" emotions is hard to say. What we _can_ say is that they are emotion
2.
▲
by
Philpax
12d ago
They didn't file the report. The model drafted a report that was not sent.
3.
▲
by
Philpax
13d ago
No, but I am suggesting that OpenAI and Anthropic are further down the RSI path than any of the other companies, as we can see from the model that solved Navier-Stokes being less than two weeks old at the time: https://openai.com
4.
▲
by
Philpax
13d ago
Are these other labs in the room with us now? No, seriously, I'm all for a multipolar world here, but he's right that the frontier is literally just those two companies at present. Google is behind. MSL is doing better, but not by
5.
▲
by
Philpax
14d ago
but like, they did The HF incident had them pwn their own cluster: https://en.wikipedia.org/wiki/2026_OpenAI_agent_cyberattacks...
6.
▲
by
Philpax
15d ago
I wish you luck on your quest to avoid all of the software on that list.
7.
▲
by
Philpax
15d ago
"Hackers" because the term in the context of "Hacker News" is broad and hard to define. Intelligence is, too, but I'm not in the business of pretending that the AI systems aren't some form of intelligent.
8.
▲
by
Philpax
15d ago
While I'm sympathetic to the sentiment, the ongoing automation of intelligence is, for better or worse, one of the most consequential things that can/will happen to "hackers" (as well as white-collar workers in general),
9.
▲
by
Philpax
16d ago
It's been around for six years; at this point, I imagine any damage it could have done has been done already.
10.
▲
by
Philpax
17d ago
Your phone does not offer more screen area on demand. That is genuinely useful for many people. Apparently, not you, but that's OK.
11.
▲
by
Philpax
18d ago
The entire point is that if he couldn't naively figure it out, no normal user would. I'm sure that he could have thought like a technical user and gotten there, but he shouldn't have needed to.
12.
▲
by
Philpax
21d ago
No, I mean we just don't know what's going on in the circuits of the model at any substantial level. We set their architecture (hyperparameters), we pump them full of data (pretraining), and we shape how they behave through exampl
13.
▲
by
Philpax
21d ago
We don't know what they do. We shape them, but our understanding of how they get to their result is comparatively minimal.
14.
▲
by
Philpax
23d ago
It was put up and then taken down. Strange things afoot.
15.
▲
OpenAI launches new Astra model amid growing scrutiny over agents' safety
(reuters.com)
3 points
by
Philpax
23d ago
|
0 comments
16.
▲
by
Philpax
23d ago
Er, o1 was also RL.
17.
▲
by
Philpax
23d ago
Actually, I would say the exact opposite. This post is full of strange and grammatically incorrect phrases, weird paragraph pacing, and unintuitive clauses: that is to say, this reads as very strongly human-written to me, and it is refreshi
18.
▲
by
Philpax
23d ago
o1 was first, and Anthropic were doing a bit of it; DeepSeek brought it to the masses, but did not invent it.
19.
▲
by
Philpax
24d ago
> Ah yes, Nexus Mods, where they'll ban you for making mods that change "Body Type A" and "Body Type B" back to "Male" and "Female", or remove mandatory pronoun selection. Oh my god please get
20.
▲
by
Philpax
24d ago
Use of Claude Code for an application used by artists. Just a fundamental misalignment of expectations.
21.
▲
by
Philpax
25d ago
Interpretation is in the eye of the beholder :-)
22.
▲
by
Philpax
25d ago
Maybe you don't! It is very possible that your problems don't actually need frontier-level artificial intelligence.
23.
▲
by
Philpax
25d ago
Ah, failed to snipe it. Here's the actual thread: https://news.ycombinator.com/item?id=49525496
24.
▲
Claude Fable 5.1
(anthropic.com)
8 points
by
Philpax
25d ago
|
2 comments
25.
▲
1Password Supports the Ethnic Cleansing of Europe
(alilleybrinker.com)
23 points
by
Philpax
26d ago
|
1 comments
26.
▲
by
Philpax
28d ago
Honestly, can't say I miss my Cursor subscription at all, and I'm surprised they're still a going concern. Why would I want to use a proprietary VSCode fork with an identity crisis when I can use literally any editor with any
27.
▲
by
Philpax
29d ago
My measurement was with MoE offloading, but there's only so much you can keep on-GPU with a 200GB quant and 48GB of VRAM. It's hard to overcome the CPU/RAM bottleneck. For what it's worth, all of my hardware was used; I
28.
▲
by
Philpax
29d ago
There are risks associated with releasing historical proprietary models that were not designed for open release: - It is trivial to extract samples of the training data that was used, which can bolster existing lawsuits/foster new ones
29.
▲
by
Philpax
29d ago
No? They were the frontier, or near it, at the time of release: https://artificialanalysis.ai/models/releases/gpt-oss-120b
30.
▲
by
Philpax
29d ago
The fastest I was able to get my Threadripper 3960X + 2x 3090s + 256GB DDR4-3200 to run a 2-bit quant of GLM-5.2 was 8 TPS. I would expect to be in seconds-per-token territory for a pure-CPU 4-bit quant.
More ›