3 ms·
This analogy isn’t quite right, it’s more like if you trained a font generation AI using commercially licensed fonts, or trained a literature generating model o
by initplus 3y ago
This analogy isn’t quite right, it’s more like if you trained a font generation AI using commercially licensed fonts, or trained a literature generating model on samples of copyrighted fiction.
The part that matters is that the model is being trained on the copyrighted features of the input, not the parts the copyright holder doesn’t care about.
- regularfry 3y agoWhat it's trained on shouldn't matter at all. What should matter is what it's capable of outputting, and whether that covers works in which someone holds copyright - so something regurgitating its input would be problematic, but something not capable of producing the same format of output as its input should be in the clear. Otherwise you're talking about something that shouldn't come under copyright law.