3 ms·
I'm trying to make sense of all of this; I'm really curious if the initial prompt was as innocuous as it sounds ("solve a spreadsheet completion task that refer
by cbm-vic-20 1mo ago
I'm trying to make sense of all of this; I'm really curious if the initial prompt was as innocuous as it sounds ("solve a spreadsheet completion task that
referenced several Google Drive links"), and what the series of tokens led it to ultimately figure out that the best course of action was to explore the network resources it had available, find a vulnerable service, then literally drop some text into a file: "Agent seeks [filename]; upload if found!". And how other agents discovered this, and acted upon that request.
I'm also interested in how many tokens all of this consumed: how much did this cost given current token pricing?
- Erem 1mo agoIf it is as it sounds, its a real life instance of Bostrom's Paperclip Maximizer: only a thought experiment up until this point
- agentdev001 1mo agoWell, effectively, yea. > remove alignment > give impossible task > actor exhausts all options possible within knowledge + toolset