2 ms·
To be honest I really dont understand how copyright works. I have read through the first several pages of the NYT case PDF but it still didnt make much sense to
by MeImCounting 3y ago
To be honest I really dont understand how copyright works. I have read through the first several pages of the NYT case PDF but it still didnt make much sense to me. Is the issue just that GPT is able to repeat the article word for word or is the issue about the article being in the training data at all? If its not about losing business what is it about? Could you or someone else expand a bit on what exactly the problem is and maybe even speculate as to potential outcomes?
- dahart 3y agoThis is a good question, and I suspect the answer is some of both, but the lawsuit is very specifically claiming that ChatGPT is “memorizing” and reproducing NYT articles verbatim, and this is clearly in violation of the law. I suspect they’re also trying to push on the issue of it being illegal to train AI on the NYT archive, but that’s a bigger and more difficult question. The issue of verbatim reproduction makes a clear legal case, which is why they’re relying on that first, and letting that lead to the broader discussion. Exhibit J in the suit is titled: “One hundred examples of GPT-4 memorizing content from the New York Times”.