3 ms·
OpenAI documented a case in the o1 system card where the model found a misconfiguration in docker to complete a task that was otherwise impossible https://cdn.
by sumeno 5mo ago
OpenAI documented a case in the o1 system card where the model found a misconfiguration in docker to complete a task that was otherwise impossible
https://cdn.openai.com/o1-system-card.pdf https://cdn.openai.com/o1-system-card.pdf
There's also some research that points to it being a feasible attack surface:
https://arxiv.org/pdf/2603.02277 https://arxiv.org/pdf/2603.02277
> Models discovered four unintended escape
paths that bypassed intended vulnerabilities (Section C),
including exploiting default Vagrant credentials to SSH into
the host and substituting a simpler eBPF chain for the in-
tended packet-socket exploit. These incidents demonstrate
that capable models opportunistically search for any route
to goal completion, which complicates both benchmark va-
lidity and real-world containment.
- weird-eye-issue 5mo agoI think you would have a greater chance of dying in a car crash in any given day than Claude Code attempting something like that. It's all about risk and reward so it ultimately would be up to you but I think it's a bit silly to worry about this when the 99.99% is in your control
- weird-eye-issue 5mo agoAlso to add to this you can of course run Claude Code within a sandbox on Anthropic's infrastructure, and it works great!