3 ms·
Courts in the relevant jurisdictions don't work on "no one really disputes." It would have to be _proven_ in a court, which involves evidence and testimony, an
by foundart 2y ago
Courts in the relevant jurisdictions don't work on "no one really disputes."
It would have to be _proven_ in a court, which involves evidence and testimony, and if the whistleblower was in a good position to provide credible testimony then his death would likely make it harder to do prove copyright violations have taken place.
- tiahura 2y agoCourts in the relevant jurisdictions don't work on "no one really disputes." It’s called a Motion for Summary Judgment.
- ghaff 2y agoI'm pretty sure any competent lawyer would stipulate that, in many/most cases, training is happening on copyrighted information. I'm also pretty sure that OpenAI is not arguing that all their training data is either licensed or they own the copyrights to. (Some companies, perhaps Adobe?, have been more conservative.) Perhaps I'm wrong. But I haven't heard that argument publicly and I would need to be convinced.
- HeatrayEnjoyer 2y agoDiscovering certain types of data were gathered and used would be much worse. Training on CNN and Netflix content = i sleep Training on private personal and corporate inboxes, medical records, and illegal content, purchased from blackhat data brokers = real shit A Kenyan data labeler famously cut ties with Openai after Openai asked them to gather CSAM content.
- BadHumans 2y agoCitation on that?
- upghost 2y agohttps://www.wsj.com/articles/chatgpt-openai-content-abusive-sexually-explicit-harassment-kenya-workers-on-human-workers-cf191483 https://www.wsj.com/articles/chatgpt-openai-content-abusive-... https://www.bigdatawire.com/2023/01/20/openai-outsourced-data-labeling-to-kenyan-workers-earning-less-than-2-per-hour-time-report/ https://www.bigdatawire.com/2023/01/20/openai-outsourced-dat... https://www.theguardian.com/technology/2023/aug/02/ai-chatbot-training-human-toll-content-moderator-meta-openai https://www.theguardian.com/technology/2023/aug/02/ai-chatbo... https://www.businessinsider.com/openai-kenyan-contract-workers-label-toxic-content-chatgpt-training-report-2023-1 https://www.businessinsider.com/openai-kenyan-contract-worke... https://www.medianama.com/2023/07/223-kenyan-workers-call-for-investigation-into-exploitation-by-openai-3/ https://www.medianama.com/2023/07/223-kenyan-workers-call-fo... They were asked to label CSAM, to clarify.
- BadHumans 2y agoGather and label are two wildly different things that change the entire context. They aren't saying go find this stuff for us, they are saying if people upload it or you find it in the data then, label it as such.
- hansvm 2y agoIt only changes who actually gathered the CSAM they asked this person to label. OpenAI definitely gathered it.