3 ms·
That's not what your Chief Research Officer, Mark Chen, says on X: "Do we use user feedback and de-identified data to improve ChatGPT and Codex in a holistic wa
by mucha 25d ago
That's not what your Chief Research Officer, Mark Chen, says on X:
"Do we use user feedback and de-identified data to improve ChatGPT and Codex in a holistic way? Yes. And so does every LLM company."
https://x.com/markchen90/status/2097400166554993041 https://x.com/markchen90/status/2097400166554993041
- vemacs 25d agoDoes OpenAI think de-identified data is no longer user data? Wild take for OpenAI and certainly not industry standard.
- sebzim4500 25d agoCan you explain what part of his post you believe is inconsistent with that quote?
- mucha 25d ago"If they opted out of training, then we definitely did not train on them." Per OpenAI's privacy policy, they use de-identified data to improve their products. From Mark Chen's comment, improving products includes improving ChatGPT and Codex in a holistic way. Improving models in a holistic way sounds a lot like training to me.
- derangedHorse 25d ago> Per OpenAI's privacy policy, they use de-identified data to improve their products That’s not inconsistent with what you responded to. They use your data unless you opt out. If the user doesn’t opt out, their de-identified data is used to improve their products.
- mucha 25d agoIt appears than you can only opt-out from having OpenAI train models on your data. There isn't an option for opting to exclude your de-identified data from being used to improve OpenAI products.
- hellohello2 25d agoAre you certain of this? I would be inclined to believe you but it would be nice to know decisively.
- hobofan 24d ago> Improving models in a holistic way sounds a lot like training to me. I think that's quite a leap. Using de-indetified data to improve the products is what everyone has been doing since the dawn of web analytics.