3 ms·
A lot of people thought that OpenAI was making this up, and I hope if you believed that, that you recalibrate your opinions of what LLM's are capable of. Worki
by empath75 2mo ago
A lot of people thought that OpenAI was making this up, and I hope if you believed that, that you recalibrate your opinions of what LLM's are capable of. Working with Fable and Opus 5 all the time, absolutely none of this surprised me capability wise, except for what seems like the long term planning capability (probably enabled by long context windows and launching subagents?)
- IAmGraydon 2mo agoVery few think they made it up. Many think they set up a situation by disabling guardrails that would inevitably end up creating a newsworthy outcome.
- signatoremo 2mo agoMany people said in the original discussion that this was more of OpenAI’s marketing than a serious issue. I counted 89 “marketing”s and 19 “stunt”s. https://news.ycombinator.com/item?id=48997548 https://news.ycombinator.com/item?id=48997548
- throwa356262 2mo agoCould still be 80% marketing. These models are trained on cyber intrusion, that's literally what ExploitGym benchmark measures. That part should not surprise anyone. But what if, say, OAI noticed the problem right away but Sam Altman recognised it would be a great PR and decided it should continue with increased compute budget?
- 0xDEAFBEAD 2mo agoWhy would you expect them to notice the problem right away? Seems likely they are doing this sort of training on a massive scale with little monitoring. "...Sam Altman recognised it would be a great PR and decided it should continue with increased compute budget?" If that's what happened, Sam should go to jail.
- estearum 2mo agoGetting more and more fun to see the "full steam ahead" people contort into more impressive shapes. Hint: If the labs making these technologies are incentivized to create or allow attacks on other services, then that is actually also a big fucking problem.
- dist-epoch 2mo agoWas HF in on it? They disabled their guardrails too, to please OpenAI? And as seen in the comments here, make many believe they are incompetent and have joke security?
- paxys 2mo agoGo read the original post. The majority of the comments were sure it was a marketing stunt.
- IAmGraydon 2mo agoGo read my post. I didn't say it wasn't. I said it wasn't "made up". In other words, it did happen. It was still a marketing stunt.