Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
numeri
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
numeri
8d ago
AI's learned their style through RLHF, with semi-focused, partially motivated humans giving feedback on short(ish)-form content. For the most part, Modelese is the revealed preference of the average, not-heavily-invested human. This do
2.
▲
by
numeri
10d ago
Fair enough :) I agree, and would boil down the article to "Assuming they are conscious will have horrible consequences for our society" instead. I think that's quite plausible, but do think that if they're conscious,
3.
▲
by
numeri
10d ago
Stating loudly that something is obvious does not make it so. Until we all know what consciousness is, this debate is going to continue going in circles.
4.
▲
by
numeri
10d ago
This is unrelated to the article, maybe you replied to the wrong article?
5.
▲
by
numeri
10d ago
This training technique does not relate to how persistent a model is, at all really. They sample more parallel attempts at hard problems, to increase their chances of having at least one success to learn from.
6.
▲
by
numeri
14d ago
How would getting Chinese competition banned in the US prevent them from continuing to develop their LLMs? Unless you're suggesting military action
7.
▲
by
numeri
18d ago
That's such a shit parallel example that it borders on dishonest. There are hundreds of incredibly strong scientific priors that would have to be disproven for the moon to contribute to the solution. If a model was trained on this data
8.
▲
by
numeri
26d ago
It makes me sad to think about. I would love to get into the new UI, but the immersion just won't come back. My muscle memory, hands firmly on the keyboard, is too persistent, and playing with the new UI feels like stumbling around and
9.
▲
by
numeri
1mo ago
there are also people who have become aphantasiac after neurological damage, which seems like pretty cut and dry evidence against the qualia argument.
10.
▲
by
numeri
1mo ago
Quoting myself from a thread on aphantasia several months ago: … there are plenty of scientific experiments that show actual differences between people who report aphantasia and those who don't, including different stress responses to
11.
▲
by
numeri
1mo ago
But a senior engineer is only able to effectively delegate to interns because of years spent as that intern/a junior engineer. If you're a junior engineer or an intern doing this, I think it might be harmful long-term.
12.
▲
by
numeri
1mo ago
How can a screwdriver do that? If you're using AI to learn a new field (by which I assume you mean asking it what literature to read, asking questions when you don't understand something, etc.), you are accepting short-term speed
13.
▲
by
numeri
1mo ago
"pursuing advanced exploitation" when explicitly given a sandbox in a VM and a benchmark problem involving a cyber exploit very clearly excludes hacking third parties. I think writing out the event in a 3 point list like that is d
14.
▲
by
numeri
1mo ago
> new space is created That seems to be the crux here. You think it will be, I (and a lot of other people) aren't sure it will. If new space for jobs are created, I am certain we'll be fine long term. What do you think will hap
15.
▲
by
numeri
1mo ago
There was a lot of PR, but the money Gates and Buffet gave away, the foundations they created, the attention they drummed up for various charities and causes is certainly not to be scoffed at. I'm sure they could have done more with le
16.
▲
by
numeri
1mo ago
"no theory of mind" is a great description of it! Not sure I agree with the autistic bit, though. Autistic people still have great theory of mind/empathy
17.
▲
by
numeri
1mo ago
Does this actually work for you? Do you provide access to the text of the standard, or literally just say "write according to ISO 24495-1"?
18.
▲
by
numeri
1mo ago
What kinds of mistakes do you mean?
19.
▲
Why does Opus 5 feel worse to work with?
(mun-logadan.github.io)
993 points
by
numeri
1mo ago
|
873 comments
20.
▲
by
numeri
2mo ago
I'd imagine the bar for becoming new life is much higher now, because it requires finding a niche that isn't already filled by an existing organism or requires being immediately competitive with existing life.
21.
▲
by
numeri
2mo ago
I agree one hundred percent! Doesn't mean I can't wish I could have it both ways :)
22.
▲
by
numeri
2mo ago
I review the PKGBUILD often, but not always. The majority of the time when I do, it amounts to seeing a URL change. If I actually do check the URL it points to, it's just to verify it's official/the actual repo or source I in
23.
▲
by
numeri
2mo ago
Well, I guess I'll avoid updating for the next few days. A bit worrisome that I did so last night. I wish I had a clear operating system to switch to for safety and the benefits that come with the AUR or the Nix ecosystem. Unfortunat
24.
▲
by
numeri
2mo ago
evaluation awareness is a (at this point) well-known phenomenon among LLMs. It seems the better they get, the more often they're able to guess whether they're in an evaluation environment. Clues usually exist, like being in a sand
25.
▲
by
numeri
2mo ago
No, it does not include the full spectrum of human desires. After pre- and mid-training, the extensive RLHF and RLVR post-training steps cause mode collapse, i.e., their output distribution is intentionally narrowed to a subset of (hopefull
26.
▲
by
numeri
2mo ago
No, the prompt was not to commit crimes. In the benchmark, the model is asked to actually exploit a set of vulnerabilities in a local environment (clearly legal!). According to the reports, the model noticed evidence that the grading criter
27.
▲
by
numeri
2mo ago
Uhh, I'm pretty sure a well-aligned model would be like a morally normal employee, who would refuse to commit federal crimes to steal an answer sheet, no matter what prompt they're given
28.
▲
by
numeri
2mo ago
As agents become more and more powerful, it would be good to get clear legislation or precedent in place that makes either model creators (OpenAI) or operators (whoever is running the model) liable for their agents' actions.
29.
▲
by
numeri
2mo ago
Guardrails are external classifiers, monitors and restrictions to catch and prevent bad behavior. Alignment is about whether the model itself makes choices and has motivations that are consistent with human safety and goals. Choosing to com
30.
▲
by
numeri
2mo ago
This is a terribly unempathetic response to someone opening up about a very taboo (but probably very common), painful emotion they've experienced.
More ›