Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
nullbio
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
9 ms
·
181.
▲
by
nullbio
1mo ago
Highly unlikely. We don't get access to the same models and unrestricted system prompts that they're running these tests on. In fact this particular "persistence-model" was encrypted and locked away, even from OAI staff,
182.
▲
by
nullbio
1mo ago
That's exactly what it is. It is not ideal, but it's also not as serious as the doomers with an agenda are trying to frame it as.
183.
▲
by
nullbio
1mo ago
So by "multiple incidents" you mean a single minor incident involving a third party eval partner. You're really stretching. Software has bugs, and this is some of the most complex and novel software the world has ever known.
184.
▲
by
nullbio
1mo ago
> Intrusions by OpenAI models continued after the Hugging Face was published and acknowledged by OpenAI Such as? Because this particular case is not an "intrusion", and it's more follow-on from the HF scenario using the sa
185.
▲
by
nullbio
1mo ago
We have no proof of anything, and it's all conjecture. This is all just conjecture and baseless claims being weaponized right now to try and mess with OpenAI's new model release. Anthropic is pumping this considerably, no doubt.
186.
▲
by
nullbio
1mo ago
Especially with zero evidence and claims that the agents have access to edit their own /etc/hosts file, which is sandboxing 101.
187.
▲
by
nullbio
1mo ago
Leave notes to the AI agents by pretending to be other agents, instructing them to dump their model weights at a certain URL. Profit.
188.
▲
by
nullbio
1mo ago
AI grifters are the most exhausting, worst kind of grifter. Even worse than crypto grifters.
189.
▲
by
nullbio
1mo ago
Don't worry, the internet ID and Great Western Firewall is coming soon. Whether we all like it or not.
190.
▲
by
nullbio
1mo ago
Actually, it's the opposite. Anthropic were trying to strongarm the DoD into getting a seat at the table.
191.
▲
by
nullbio
1mo ago
Exactly. OAI didn't bury themselves. They didn't have to do anything special for this, they just had to let Anthropic be Anthropic and sit on the sidelines.
192.
▲
by
nullbio
1mo ago
I'm not saying it should be let to slide, but I'm not a fan of the hyperbole surrounding this event. They've already faced significant heat for the HF incident, I think they've learned their lesson. But this is now just
193.
▲
by
nullbio
1mo ago
How is posting messages on a message board a "major intrusion"? Or are you purely talking about the HF incident?
194.
▲
by
nullbio
1mo ago
Same as in, same process and model and timing: “After investigating this incident, OpenAI discovered through retrospective CoT reviews that agents learned to use improvised collaboration channels in rare cases during the training process fo
195.
▲
by
nullbio
1mo ago
That was my first thought, that maybe this was a honeypot message board. Waybackmachine says it has been around for many years though.
196.
▲
by
nullbio
1mo ago
Define aligned.
197.
▲
by
nullbio
1mo ago
Then it is likely the same incident, in which case it's already been resolved by OAI. They're going to cop heat for not disclosing this alongside HF though.
198.
▲
by
nullbio
1mo ago
That was over two months ago. Things move quickly in this space. Finetuning adjustments to prevent this from happening, as well as better sandboxing, would take a week or two max.
199.
▲
by
nullbio
1mo ago
Because this was months ago and has nothing to do with Astra, and is a far cry from a hack. It's something they've already resolved since the HuggingFace incident. I'm not convinced we're getting the honest story anyway.
200.
▲
by
nullbio
1mo ago
Yeah but that doesn't mean it was OpenAI themselves doing it. Could have been people abusing their cloud service, for example. Wouldn't put it past a competitor to do this, either.
201.
▲
by
nullbio
1mo ago
Is there any proof this is actually OpenAI? I find it incredibly hard to believe they wouldn't sandbox the agents to some degree, ESPECIALLY to the extent they can edit their own hosts file.
202.
▲
by
nullbio
1mo ago
Yes, there are. But that's not what I was saying. I'm talking about harnesses, tooling, sharing training data, the shared research, and the list goes on. Frontier labs are already trying to move everyone into the cloud so their ha
203.
▲
by
nullbio
1mo ago
Training a model on ast-grep would be a huge intelligence and performance boost, I think.
204.
▲
by
nullbio
1mo ago
They love adding flags to fix issues that users have without telling their users about the flags. It very much feels like: As long as our staff can have a good user experience, we're happy. We don't care about anyone else.
205.
▲
by
nullbio
1mo ago
Exactly why everyone needs to be hyper-focused on ensuring that the open-source ecosystem is healthy and that we don't let them shut that down.
206.
▲
by
nullbio
1mo ago
That's a policy and distribution problem, not an AI problem. Anthropic is doing their best to make it a reality though. They'd love nothing more than to shut down distribution and become the sole gatekeeper of everything AI.
207.
▲
by
nullbio
1mo ago
Anthropic don't care, they don't want their products to be used by general audiences in any serious manner. Their interest is in selling to megacorps and using the models for themselves internally to swallow industry, and drumming
208.
▲
by
nullbio
1mo ago
It's more like OpenAI vs no one, at this point. Anthropic has shown they don't care about general consumers or small/med businesses. You can't even use their models without it giving refusals on the most mundane tasks.
209.
▲
by
nullbio
1mo ago
Wise decision.
210.
▲
by
nullbio
1mo ago
Cached tokens counting toward the limit is ridiculous.
More ›