Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
mofeien
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
33 ms
·
31.
▲
by
mofeien
2mo ago
Possible solution/mitigation: have a mandatory "sender" field? Then every inciting message can at least be attributed to the government instead of just a disembodied voice from your phone. That my help people contextualize an
32.
▲
by
mofeien
2mo ago
freudian slop?
33.
▲
by
mofeien
2mo ago
From TFA: It did succeed in the "accidentally impossible" task, but not at all in the way the problem-setters intended, and rather... at all costs?! And it wouldn't really matter whether it stopped afterwards, I think. At suf
34.
▲
by
mofeien
2mo ago
I don't think air gapping will work: even human security researchers recovered a 378-bit key from a Samsung Galaxy S8 through a power LED of a speaker two devices away. And accessing memory in a specific sequence can generate radio sig
35.
▲
by
mofeien
2mo ago
To what kind of goal that an ASI might decide to pursue would "a quarter billion happy, healthy, free people" be the most efficient solution to?
36.
▲
by
mofeien
2mo ago
That's an application of the "Tragedy of the commons": One person stopping to push the frontier just means that this person will lose their job, the frontier will be pushed by the others. The solution is to push for a binding
37.
▲
by
mofeien
2mo ago
It means that they request the structures needed to navigate the event horizon of recursive self improvement be put into place. Probably: International agreements, verifiability of an AI pause on the level of graphic cards as well as throug
38.
▲
by
mofeien
3mo ago
Scoring well on the first clears the background knowledge for practical work. The second test from the quote above was about synthesizing pathogens and it synthesized plasmids for 2 out of the 10 pathogens.
39.
▲
by
mofeien
3mo ago
Transcript, four prompts: https://chatgpt.com/share/6a60b2eb-0b64-83ee-9c76-7931ca1de0...
40.
▲
by
mofeien
3mo ago
From the Fable 5 System card: > Results > On the VCT multimodal virology evaluation, Mythos 5 scored 0.56, well above the expert baseline of 0.221 and nearly matching that of Mythos Preview (0.57). This represents an improvement over
41.
▲
by
mofeien
4mo ago
https://status.claude.com/ You can also just subscribe here for updates on fable 5 status
42.
▲
by
mofeien
4mo ago
The top comment in the very discussion you linked on that AISLE blog has a strong rebuttal to that blog post...
43.
▲
by
mofeien
4mo ago
I would assume that shortly after, the solar system will be hyper optimized as well, then the milky way, then the local cluster, and so on. Everything will be close to optimal afterwords, and I sure hope we will have specified the target fu
44.
▲
by
mofeien
4mo ago
Or agree on finding ways to promote peaceful use of nuclear energy. This has been done, there are thousands of people working on it around the globe and 180+ member states of the IAEA. It's not easy, there have been close calls. And co
45.
▲
by
mofeien
4mo ago
The regulation that is being argued for here is against pushing the frontier. Entering the market with say a new speech to text model is not subject to such regulation. What's needed is something qualitatively different from entry barr
46.
▲
by
mofeien
4mo ago
> If it were possible to effectively slow the development of this technology to give ourselves more time to deal with its immense implications, we think that would likely be a good thing Even Anthropic wants to Pause AI now. There must r
47.
▲
by
mofeien
4mo ago
Jack Clark, co-founder of Anthropic said the following at an Oxford lecture last week ([0], at around 10 and 12 mins): "It's a technology that we do not fully understand because it's more grown than made. And it is a te
48.
▲
by
mofeien
4mo ago
Not all training data is human generated, and it's also not clear that being ridiculously good at interpolating between data points (whatever that means) will not lead to superhuman capabilities.
49.
▲
by
mofeien
5mo ago
- Its goal: X - (Logic) => its subgoal: Not be turned off because that's a prerequisite to be able to do X - (Logic) => Eliminate humans with their opaque and somewhat unpredictable minds to reduce chance of harm to it from 0.01%
50.
▲
by
mofeien
5mo ago
The goal as stated on the extension page is to improve the readability of texts by replacing :, *, _ forms. So some customizability to the user's wishes would be quite nice. My calculus textbook (Königsberger, 2004) in university used
51.
▲
by
mofeien
5mo ago
To prevent accusations of "masculinism" or sexism and to have a stronger case on having the goal to improve readability the add-on could include an option (or even make it default) to replace by generic feminine instead.
52.
▲
by
mofeien
5mo ago
There was no nuclear weapon used in warfare anymore since WW II. I think the regulation and oversight worked incredibly well over the past 70-80 years, despite the game-theoretic challenge you mention.
53.
▲
by
mofeien
5mo ago
Given that his reason for saying GPT-2 was too dangerous to release was that the world needed more time to prepare for the effects of this technology, and given that the following models were basically scaled-up versions of it and killed so
54.
▲
by
mofeien
5mo ago
"The race to build smarter-than-human AI is a race with no winners." And specifically about the point on China, several people in power in China have also expressed the need to regulate AI and put international structures of gover
55.
▲
by
mofeien
5mo ago
That highlights how important ceiling construction regulations are. I would assume that right now your breakfast sandwich is more highly regulated than LLMs. And these are the things that make decisions spanning from database maintenance he
56.
▲
by
mofeien
6mo ago
These results were based on "a trivial snippet from the OWASP benchmark". In the section "caveats and limitations" they state that sonnet 4.6 and opus 4.6 now pass. And they decided to base the false positive examination
57.
▲
by
mofeien
6mo ago
... or maybe when you see them triggered or exploited reproducibly, then the underlying bug will also be pretty easy to discover. But at that point, it's already too late. :) I really like your original point, I never thought about it
58.
▲
Anthropic has just built an AI that could take down the internet
(pauseai.substack.com)
3 points
by
mofeien
6mo ago
|
0 comments
59.
▲
by
mofeien
6mo ago
I can think of several possible messy outcomes that would be able to directly affect me, not all mutually exclusive: - Job loss by me being replaced by an AI or by somebody using an AI. Or by an AI using an AI. - Resulting societal instabil
60.
▲
by
mofeien
6mo ago
I am freaking out. The world is going to get very messy extremely quickly in one or two further jumps in capability like this.
More ›