26 ms·
I think it's wise to take the concerns of the creative community seriously - after all their "labor of love" [1] matters immensely, without it LLMs are useless.
by al_be_back 3y ago
I think it's wise to take the concerns of the creative community seriously - after all their "labor of love" [1] matters immensely, without it LLMs are useless.
matters not how much the coder "loved" the project, or did yoga, or that they've not made money for years, after all most book authors aren't exactly raking in the money either.
also like many thinks in life, some tools/projects/startups etc just stop being needed/used and new ones/competitors take over. there's nothing to say that since tool X is using A.I. therefore it has to be adopted by one and all; smiles all around.
google has 'right to be forgotten', also looking into 'machine unlearning' and it's common for platforms to honor user's request to remove their data / close their account.
[1] From OP: destroying what had been a clear labor of love and a useful project
- tavavex 3y agoThe thing with this project is that it had no conceivable way of threatening the original authors, financially or otherwise. The analyses it produced weren't replacements of the original works, and served as a writing tool, not as something that generated new content. In this case, the shutdown seemed to hinge entirely on the irrational fear of anything with "AI" written on it and an unshakeable conviction that this clearly transformative use case wasn't fair use.
- al_be_back 3y ago>> had no conceivable way of threatening the original authors, financially or otherwise how so? what's inconceivable about it? >> seemed to hinge entirely on the irrational fear how are the authors "fearing without reason" or "illogically fearing" ?
- tavavex 3y ago> how so? what's inconceivable about it? Authors make money through sales of their work. This tool was a writing aid that analyzed text and included some copyrighted works in its dataset. There was no way of retrieving these books in full, and the excerpts that were allegedly shown to the users were used in an analytical context, unlike the original works. So, this website couldn't replace ownership of the actual book for readers, and had no capacity to hurt their sales. Basically, the way the service used this data should be transformative enough as to not have any impact on the authors. > how are the authors "fearing without reason" or "illogically fearing" ? I called it "fear" because there was no strong argument on the authors' side as to why this tool is bad. I called it "illogical" because I think that it's no coincidence that this controversy only came up now, in 2023. Back in 2017 and onward, the existence of this tool didn't appear to generate pushback. My pet theory is that in 2023, now that we have good generative AI, the advancements have spawned an entire subset of people that view anything "AI" as inherently tainted and immoral. The original complaints seem to lack understanding of this tool and multiple people have conflated it with generative AI, despite it having nothing to do with that.
- al_be_back 3y ago>> Basically, the way the service used this data it's not about the past (2017...), the authors are concerned about how the dataset could be or is likely to be used from now on. many tech projects these days have or are thinking about integrating third-party A.I. providers in their services, either to harness the power of their large datasets or their large user-base. I think it's great if authors/users opt-in to this, but likewise I agree with those that want out (opt-out). >> I called it "fear" because there was no strong argument on the authors' side as to why this tool is bad their argument doesn't have to be a peer-reviewed journal, it suffices to say "i don't want my books in your dataset" >> I called it "illogical" because I think that it's no coincidence that this controversy only came up now, in 2023. Back in 2017 and onward, the existence of this tool didn't appear to generate pushback ... 6 years have passed since 2017, life moves on, it's natural for things to change e.g: the project's code, updating servers, partnerships, emergence of third-party tools/libs/services etc etc.
- tavavex 3y ago> the authors are concerned about how the dataset could be or is likely to be used from now on. The service in question wasn't introducing anything new that'd appear to justify all the recent pushback. It feels like you continue generalizing your statements, while the discussion topic is about what made Prosecraft specifically so preposterous that it warranted the outrage. > their argument doesn't have to be a peer-reviewed journal, it suffices to say "i don't want my books in your dataset" It's kind of a blunt statement, but why should they have a say? For example, say I create a website where I publish technical analysis of famous literary works, including basic statistics about a book and a review. Should the authors be able to just take that down? This use is legally protected (as is creating a dataset), so allowing authors to restrict this use seems as arbitrary as allowing them to say that no person can ever bring their books into the country of Moldova or that no one over the age of 50 may read it. > it's natural for things to change e.g: the project's code, updating servers, partnerships, emergence of third-party tools/libs/services etc And yet, in this specific situation, all of this is conjecture. Nothing about the project changed in some significant way in 2023 that would warrant this. Further proving it is that the people that are against Prosecraft don't seem to bring up any specific changes or reasons for their stance, only that it is "AI".