3 ms·
> Releasing the trained model does not make it open source. This is a straw man argument as Meta has released the model weights and not just the trained model.
by menzoic 3y ago
> Releasing the trained model does not make it open source.
This is a straw man argument as Meta has released the model weights and not just the trained model. When it comes to models the weights are the source.
- swatcoder 3y agoIt may be a difference in terminology, but it's definitely no straw man. There's a philosophy of open source based on recipients being able to make any small and large modifications to the source material as they see fit. Just being able to build new projects with something or read what it does not qualify in that context -- it requires the freedom to rework the thing itself to suit your own needs and ends. In the case of these models, that philosophy would have the source training data and the training algorithms themselves available, editable, verifiable, and repeatable. Ignoring the financial challenges, a truly open source model would let you remove, add, edit, or reclassify a sample in the original training data and then build new weights of their won. The released weights are not that kind of open source, and that's what these arguments are generally making the a case for. There are still plenty of grounds from which to critique that argument if you wanted to, but there's no straw man involved in it.
- andy99 3y agoI don't want to support the guy you replied to, he doesn't understand what open source means, but I belive and have argued it is possible to exercise the freedoms that are part of open source software without access to the training data. See https://www.marble.onl/posts/considerations_for_copyrighting_AI.html https://www.marble.onl/posts/considerations_for_copyrighting...
- cactusplant7374 3y agoDid all of these organizations acquire training data in a legal manner? Seems like we are tip toeing around the real issue here.
- menzoic 3y ago>There's a philosophy of open source based on recipients being able to make any small and large modifications to the source material as they see fit. The reason I see it as a strawman is because the weights are the source of the AI model and they do allow you to modify the model however you please. The data is not owned by Meta, they can’t release it and it’s not required to use or modify the model. This is like wanting a team to release the research that informed product development with their source. The data and training code are not the same as the AI model itself. This is not the same as code that produces a binary, this is more like code that produces another form of code. If you try to replicate the model you won’t be able to because the process is not deterministic. There’s no reason to replicate the model, you would fine tune it instead.
- deleted 3y ago[deleted]