5 ms·
> Is it unfair for you to create content/products/etc after you have read and learned from various sources on the internet, potentially depriving them of clicks
by _qzu4 4y ago
> Is it unfair for you to create content/products/etc after you have read and learned from various sources on the internet, potentially depriving them of clicks/income?
Because it's false equivalence? ChatGPT isn't a human being. It's a product that is built upon data from other sources.
The question is if this data is legal to scrape, which it is: Web scraping is legal, US appeals court reaffirms [https://news.ycombinator.com/item?id=31075396 https://news.ycombinator.com/item?id=31075396].
As long as the content is not copyrighted and it's not regurgitating the exact same content, then it should be okay.
- jMyles 4y ago> The question is if this data is legal to scrape ...it is? I didn't see that question raised in OP's text at all. What do legacy human legalities have to do with how AI will behave? > Because it's false equivalence? ChatGPT isn't a human being. Is this important? What is so special about human learning that it puts it in a morally distinct category from the learning that our successors will do? It sounds like OP is concerned with the ad-driven model of income on the internet, and whether it requires breaking in order for AI to both thrive and be fair.
- venv 4y ago>Is this important? Well yes, it's the whole crux of the matter. Laws govern human behaviour. As of 2023, only living beings have agency. If I shoot someone with a gun, the criminal is me and not the gun. Being a deterministic piece of silicon, a computer is perfectly equivalent. Sure, it is important to start a discussion of potential nonhuman sentience in the future, but these AI models are not unlike any previous software in legal issues. It's bizarre to me how many people are missing this.
- nowherebeen 4y ago> It's bizarre to me how many people are missing this. Very much this. I am too tired right now to engage with other responders, but thank you for articulating precisely the point I want to make.
- jMyles 4y ago> these AI models are not unlike any previous software in legal issues Agreed. However, the previous 'legal issues' related to software and the emergence of the internet are also difficult to take seriously when considered on anything but extremely short time scales. Every time we swirl around this topic, we arrive at the same stumbles which the legacy legal system refuses to address: * If something happening on the internet is illegal, _where_ is it illegal? Different jurisdictions recognize different jurisdictional notions - they can't even agree on whose laws apply where. If you declare something to be illegal in your house, does that give it the force of law on the internet? Of course not. Yet, the internet doesn't recognize the US state any more than it does your household. It seamlessly routes around the "laws" of both. * The "laws" that the internet is bound to follow are the fundamental forces of physics. There is no - and can be no - formal in-band way for software to be bound to the laws of men, because signals do not obey borders. The only way to enforce these "laws" are out-of-band violence. * States continuously, and without exception, find themselves at a disadvantage when they make the futile effort to stem the evolution of the internet. For example, only 30 years ago (a tiny spec in evolutionary time scales), the US state gave non-trivial consideration to banning HTTPS. I understand that people sometimes follow laws. But they also often don't. The internet has already formed robust immunity against human laws. Whatever human laws are, they are not the crux of anything related to evolution of software. They are already routinely cast aside when necessary, and are very clearly headed for total irrelevance.
- Retr0id 4y agoBeing allowed to scrape something does not absolve you of all intellectual property, copyright, moral, etc. issues arising from subsequent use of the scraped data.
- dawsoneliasen 4y agoExactly, besides, the question isn’t about legality, it’s about what the law should be, I think. The question isn’t whether it’s legal, the question is whether we need to change the law in response to technology.
- faktory 4y agoChatGPT isn't doing the scraping, humans are. And humans are using computers to both read the article and create content or to scrape it. So not it's not a false equivalence.
- anileated 4y agoThere’s a reason scraping is a legally grey area. > Web scraping is legal, US appeals court reaffirms First, the case is not closed. [0] Second, to draw an analogy, you can use scraping in the same way you can use a computer: for legal purposes. That is, you cannot use scraping to violate copyright, just as you cannot use a computer to violate copyright. The following being my conjecture (IANAL), there is fair use and there is copyright violation, and scraping can be used for either—it does not automatically make you a criminal, but neither is it automatically OK. If what you do is demonstrably fair use presumably you’d be fine; but OpenAI with its products cannot prove fair use in principle (and arguably the use stops being fair already at the point where it compiles works with intent to profit). [0] https://news.ycombinator.com/item?id=31079231 https://news.ycombinator.com/item?id=31079231
- faktory 4y agoYes but that's a technical issue. I took the parent as making a philosophical point and responded in that spirit.
- williamcotton 4y agoWouldn’t it be nice if the people on these forums were not ignorant of both philosophy or the legal system before diving into incoherent conversations about both at the same time where the main thrust is the emotions they have about these tools?
- anileated 4y agoOne can dream.
- 4y ago
- bilsbie 4y ago> It's a product that is built upon data from other sources. To be fair, so are you.
- anonymouskimmer 4y agoCheck me on this because I'm not a software person: When a person "scrapes" a website by clicking through the link it registers as a hit on the website and, without filters being turned on, triggers the various ad impressions and other cookies. Also if the person needs that information again odds are they'll click on a bookmark or a search link and repeat the impression process all over again. When an AI scrapes the web it does so once, and possibly in a manner designed to not trigger any ads or cookies (unless that's the purpose of the scrape). It's more equivalent to a person hitting up the website through an archive link.