39 ms·
Given that it's public, yes, gitlab as well as every other model on earth will be trained from it. Is it right? I don't know. Can we stop it? Absolutely not.
by arwineap 2y ago
Given that it's public, yes, gitlab as well as every other model on earth will be trained from it.
Is it right? I don't know. Can we stop it? Absolutely not.
Honestly, I find it comforting that gitlab is at least being straightforward about it
- cbsmith 2y ago> Is it right? I don't know. Can we stop it? Absolutely not. Depends on what you mean by "stop it". Patents are public, but it's hard to employ them for commercial purpose without an agreement from the owner.
- bondarchuk 2y agoI would not call that straightforward. Being straightforward would mean including something like "we & our vendors might (or will) train generative AI models based on public data", instead of letting us infer that for ourselves.
- freedomben 2y agoStraightforward? I very much disagree. You could maybe convince me that it is straightforward when measured against only corporate speak, but if not looked at as corporate speak it is very much not straightforward. Straightforward would be an explicit "Yes, we train from public code regardless of the license"
- danielmarkbruce 2y agoThe only statement they can make which is accurate is that they will not train on private data. They might train on your data if it's public. If your code is public, and it doesn't get pulled into their dataset, they won't train on it.