13 ms·
Copyright issues don't seem to be addressed by any large language model provider. If an LLM is trained on GPL code then that code has become an intrinsic part
by breve 4mo ago
Copyright issues don't seem to be addressed by any large language model provider.
If an LLM is trained on GPL code then that code has become an intrinsic part of the model (because if it hasn't then what was the value of training on it). So shouldn't that model now also be licensed GPL?
And how do I know the LLM output is not reproducing substantial chunks of GPL'd code, making my code GPL?
- imglorp 4mo agoMaybe this, but multiply by N licenses. Any given output may have ideas from all of them. Law is probably going to take a while to catch up here.
- Ekaros 4mo agoOr alternatively. LLM is not human. Non human generated content has no copy right protection. Meaning all generative model output is automatically public domain.
- lefra 4mo agoI don't understand this argument. When someone uses a tool to create something, they are the copyright holder. Why would it be different when the tool is an LLM?
- olsondv 4mo agoGithub copilot has filters for enterprise that remove the GPL code before it gets returned. At least that’s how my company has been covering itself.