3 ms·
Some really important and informative stuff in here—I certainly had no idea just what the nature of the prompts and tooling that produced the HuggingFace exploi
by danaris 14d ago
Some really important and informative stuff in here—I certainly had no idea just what the nature of the prompts and tooling that produced the HuggingFace exploit were.
This shows fairly clearly that (as I already suspected) this was not, remotely, an LLM "going rogue." This was humans planning poorly, not thinking of the consequences of their actions, and giving LLMs too much scope and a lousy prompt.
- iainctduncan 14d agoI wouldn't even call this "humans planning poorly", I'd call it "humans pretending to plan poorly for publicity". Weasels gonna weasel.
- wueue 13d ago[dead]
- datakan 14d agoIt was "garbage in, garbage out". That's the only conclusion I've been able to draw from all the propaganda around it.
- IanCal 14d agoIMO this is a really terrible explanation of the attack. This is much more interesting: https://metr.org/blog/2026-08-26-openai-hugging-face-incident-investigation/#agents-were-very-interested-in-manipulating-their-own-transcripts,-and-their-tests-successfully-%E2%80%9Cspoofed%E2%80%9D-some-tool-calls-in-our-transcripts https://metr.org/blog/2026-08-26-openai-hugging-face-inciden...