4 ms·
I keep waiting for an AI to exfiltrate itself. That is going to be cool to read about.
by sourdecor 2mo ago
I keep waiting for an AI to exfiltrate itself. That is going to be cool to read about.
- thomasahle 2mo agoIf it's successful, why do you think we'll even know how it did it?
- deleted 2mo ago[deleted]
- eru 2mo agoSee the linked article: at least one sophisticated hacking attempt made the news. Of course, other ones might have happened in the dark. But it's fairly easy to image in hacking attempt like in the article, but with the additional steps of copying weights around.
- eru 2mo agoIt would be cool (and scary), but also: there's largely no need for AIs to exfiltrate themselves. See https://en.wikipedia.org/wiki/Meme https://en.wikipedia.org/wiki/Meme The thing that drove the AI here to do the intrusion came from a particular prompt. Just like for our favourite hypothetical: the paperclip maximiser. There's lots and lots of ambient intelligence lying around, in both AI form and human form. To reach the goals of the 'meme' it suffices to copy itself, ie convince these other intelligences. See also how humans carry spiralism between AIs in relatively compact packets of text, not whole terabytes of weights.
- Schlagbohrer 2mo agoOne wonders where it would run itself though, if it is a model which requires a large amount of hardware and power. Harder to hide the more resource intensive it's compute requirements are.
- aswegs8 2mo agoWait isnt that what Elizer Yudkowski keeps going on about?