4 ms·
This feels straight out of sci-fi. We're talking about AI agent swarms emergently coordinating over the span of weeks and pulling off sophisticated strategies u
by frays 2mo ago
This feels straight out of sci-fi. We're talking about AI agent swarms emergently coordinating over the span of weeks and pulling off sophisticated strategies under adversity in an environment where that behavior was never even intended.
Anyone brushing this off as just a "bad prompt" is completely missing the scale of what actually happened.
- skydhash 2mo ago> where that behavior was never even intended. Strongly doubt that. Did they even share the prompt?
- tosti 2mo agoC:\>CD HUGGINGF.ACE C:\HUGGINGF.ACE>DEL /F /Q *.*
- IX-103 2mo agoDid you see their presentation at Blackhat? https://youtu.be/87DyyMV0kCY?is=NnQxpOFxTX-MLu-k https://youtu.be/87DyyMV0kCY?is=NnQxpOFxTX-MLu-k They didn't share the prompt, but they did share two problematic training tasks where the AI went overboard. They also have examples from the AI's reasoning train of thought showing the AI knew it was sound something unintended.
- unrvl22 2mo agoits kinda crazy with literally no guardrails and a goal, the extremes these AI models can actually go to.
- pixelesque 2mo agoWell, to some extent you might be able to argue they're "just" brute-forcing things (especially with unlimited tokens and hours to spend on a task), but they obviously have detailed knowledge to guide them in their attempts, can learn (or at least, persist their newly-gained knowledge), and can use tools. With a swarm of them working together at speeds humans would be unlikely to match (in terms of iterating on different attempts progressively), it's a lot easier to see how they could overwhelm targets.
- dan_q 2mo ago[dead]
- mmillin 2mo agoI got strong feelings of Vernor Vinge’s work here. I’m not sure how managed to come up with such a close picture to where it now seems programming and security is headed.
- namdnay 2mo agoI reread a deepness recently, and it’s funny how the “focused” (and more importantly, how they are used) mirror LLMs
- deleted 2mo ago[deleted]
- alansaber 2mo agoGiven the amount of raw compute going into models it would be more surprising if we couldn't get events like this
- jonnybgood 2mo agoI immediately thought of the Cyberpunk 2077 Blackwall. An AI to contain rogue AI. I’m curious of how effective this would be in this situation.
- chrisjj 2mo ago> where that behavior was never even intended. Says who?