13 ms·
> How will you store my data and who can access it? > The content covered by the court order is stored separately in a secure system. It’s protected under lega
by _jab 1y ago
> How will you store my data and who can access it?
> The content covered by the court order is stored separately in a secure system. It’s protected under legal hold, meaning it can’t be accessed or used for purposes other than meeting legal obligations.
> Only a small, audited OpenAI legal and security team would be able to access this data as necessary to comply with our legal obligations.
So, by OpenAI's own admission, they are taking abundant and presumably effective steps to protect user privacy here? In the unlikely event that this data did somehow leak, I'd personally be blaming OpenAI, not the NYT.
Some of the other language in this post, like repeatedly calling the lawsuit "baseless", really makes this just read like an unconvincing attempt at a spin piece. Nothing to see here.
- lxgr 1y agoIf the stored data is found to be relevant to the lawsuit during discovery, it becomes available to at least both parties involved and the court, as far as I understand.
- sashank_1509 1y agoObviously openAI’s point of view will be their point of view. They are going to call this lawsuit baseless, they would not be fighting it or else.
- ivape 1y agoTo me it's pretty clear the way this will happen. You will need to buy additional credits or subscriptions through these LLMs that feedback payment to things like NYT and book publishers. It's all stolen. I don't even want to hear it. This company doesn't want to pay up and willing to let user's privacy hang in the balance to draw the case out until they get sure footing with their device launches or the like (or additional markets like enterprise, etc).
- fallingknife 1y agoCopyright is pretty narrowly tailored to verbatim reproduction of content so I doubt they will have to pay anything.
- tiahura 1y agoincorrect. copyright applies to derived works.
- vel0city 1y agoEven then, it's possible to prompt the model to exactly reproduce the copyrighted works.
- fallingknife 1y agoPlease show me one of these prompts
- vel0city 1y agoNYT has examples in their legal complaint. See page 30. https://www.scribd.com/document/695189742/NYT-v-OpenAI https://www.scribd.com/document/695189742/NYT-v-OpenAI
- Workaccount2 1y ago> It's all stolen. LLMs are not massive archives of data. The big models are a few TB in size. No one is forgoing a NYT subscription because they can ask ChatGPT to print out NYT news stories.
- edbaskerville 1y agoRegardless of the representation, some people are replacing news consumption generally with answers from ChatGPT.
- tptacek 1y agoNo, there is a whole news cycle about how chats you delete aren't actually being deleted because of a lawsuit, they essentially have to respond. It's not an attempt to spin the lawsuit; it's about reassuring their customers.
- VanTheBrand 1y agoThe part where they go out of the way to call the lawsuit baseless is spin though, and mixing that with this messaging does exactly that, presents a mixed message. The NYT lawsuit is objectively not baseless. OpenAI did train on the Times and chat gpt does output information from that training. That’s the basis of the lawsuit. NYT may lose, this could end up being considered fair use, it might ultimately be a flimsy basis for a lawsuit, but to say it’s baseless (and with nothing to back that up) is spin and makes this message less reassuring.
- tptacek 1y agoNo, it's not. It's absolutely standard corporate communications. If they're fighting the lawsuit, that is essentially the only thing they can say about it. Ford Motor Company would say the same thing (well, they'd probably say "meritless and frivolous").
- hiddencost 1y ago> So, by OpenAI's own admission, they are taking abundant and presumably effective steps to protect user privacy here? In the unlikely event that this data did somehow leak, I'd personally be blaming OpenAI, not the NYT. I am not an Open AI stan, but this needs to be responded to. The first principle of information security is that all systems can be compromised and the only way to secure data is to not retain it. This is like saying "well I know they didn't want to go sky diving but we forced them to go sky diving and they died because they had a stroke mid air, it's their fault they died.". Anyone who makes promises about data security is at best incompetent and at worst dishonest.
- JohnKemeny 1y ago> Anyone who makes promises about data security is at best incompetent and at worst dishonest. Shouldn't that be "at best dishonest and at worst incompetent"? I mean, would you rather be a competent person telling a lie or an incompetent person believing you're competent?
- HPsquared 1y agoAn incompetent but honest person is more likely to accept correction and respond to feedback generally.
- nhecker 1y agoData is a toxic asset. -- https://www.schneier.com/essays/archives/2016/03/data_is_a_toxic_asse.html https://www.schneier.com/essays/archives/2016/03/data_is_a_t...
- pritambarhate 1y agoMay be because you are not OpenAI user. I am. I find it useful and I pay for it. I don't want my data to be retained beyond what's promised in the Terms of Use and Privacy Policy. I don't think the Judge is equipped to handle this case if they don't understand how their order jeopardies the privacy of millions of users worldwide who don't even care about NYT's content or bypassing their paywalls.
- mmooss 1y ago> who don't even care about NYT's content or bypassing their paywalls. Whether or not you care is not relevant, and is usually the case for customers. If a drug company resold an expensive cancer drug without IP, you might say 'their order jeopardies the health of millions of users worldwide who don't even care about Drug Co's IP. If the NYT is right - I can only guess - then you are benefitting from the NYT IP. Why should you get that without their consent and for free - because you don't care? > (jeapordizes) ... is a strong word. I don't see much risk - the NYT isn't going to de-anonymize users and report on them, or sell the data (which probably would be illegal). They want to see if their content is being used.
- conartist6 1y agoYou live on a pirate ship. You have no right to ignore the ethics and law of that just because you could be hurt in conflict related to piracy
- DrillShopper 1y agoThe OpenAI Privacy Policy specifically allows them to keep data as required by law.