4 ms·
If the data is opensource on github, then in my opinion it should be fair game.
by prism56 5mo ago
If the data is opensource on github, then in my opinion it should be fair game.
- ozgrakkurt 5mo agoIMO this is unfair for GPL or similarly licensed code. Seems ok for MIT like licensed code though
- ForHackernews 5mo agoIt's totally fair to use GPL code, it just means all the models built by Anthropic, OpenAI, etc. using GPL-licensed source are themselves bound by the GPL. Plus, any works created downstream using those AI tools. We're on the verge of a golden age of software as soon as someone finds a court with courage.
- duskdozer 5mo agoAh, you have much more faith in the legal system than I do. It's nice to dream, though.
- edg5000 5mo agoI think AI will create an open source dark age. Gradually, we'll see a lot less new good open source code. A gradual shift back to the proprietary world. Simmilar to the 1950-1990 period.
- singpolyma3 5mo agoWhy would giving more people software freedom and the ability to reverse engineer nonfree code result in a dark age?
- singpolyma3 5mo agoThere's no difference. Either you need to follow the license or you don't. MIT has requirements still.
- notrealyme123 5mo agoThings being public should not be enough. just because someone leaked your medical information to the public via a data breach should not make it fair game. There should be some rules.
- prism56 5mo agoI feel that's a flase dichotomy. The code visible on github is freely available for anyone to read and learn from.
- notrealyme123 5mo agoSo would be your leaked medical record. The point is not that this situation seems absurd. The point is that we need some point where we say whats ok or not. And by ignoring licensing of public code already we moved it closer to the worse end of the spectrum
- prism56 5mo agoI feel that's a false dichotomy. The code on github is freely available for people to read and learn from, leaked medical data isn't.
- singpolyma3 5mo agoThere are rules. I believe that search engine indexing follows these rules and that so called "training" is search engine indexing. But a court may differ in the future.
- driverdan 5mo agoThe data is not open source. They have open weights but the source data is never open.