5 ms·
a subscription doesn't give you an automatic escape hatch out of copyright law. here's their ToS, which is pretty clear about what you cannot do: https://help.
by test098 3y ago
a subscription doesn't give you an automatic escape hatch out of copyright law.
here's their ToS, which is pretty clear about what you cannot do: https://help.nytimes.com/hc/en-us/articles/115014893428-Terms-of-Service#4 https://help.nytimes.com/hc/en-us/articles/115014893428-Term... (relevant parts below)
Without NYT’s prior written consent, you shall not:
...
(2) use robots, spiders, scripts, service, software or any manual or automatic device, tool, or process designed to data mine or scrape the Content, data or information from the Services, or otherwise use, access, or collect the Content, data or information from the Services using automated means;
(3) use the Content for the development of any software program, including, but not limited to, training a machine learning or artificial intelligence (AI) system.
...
(5) cache or archive the Content (except for a public search engine’s use of spiders for creating search indices);
- jasomill 3y agoFull text New York Times articles are available through subscription services other than nytimes.com. As an example, my local library offers full-text NYT articles through both nytimes.com and ProQuest. Notably, ProQuest's terms only explicitly ban scraping metadata and developing software or services that "compete or interfere" with ProQuest products: https://about.proquest.com/en/about/terms-and-conditions https://about.proquest.com/en/about/terms-and-conditions
- test098 3y agoand your local library and ProQuest are also bound by the same laws, even if they have an existing licensing agreement. from the ToS you just linked (note use of the terms "licensor" and "third party", which would be the NYT in this case): > Restrictions. Except as expressly permitted above, Customer and its Authorized Users shall not: > Remove any copyright and other proprietary notices placed upon the Service or any materials retrieved from the Service by ProQuest or its licensors; > Perform automated searches against ProQuest’s systems (except for non-burdensome federated search services), including automated “bots,” link checkers or other scripts; > Provide access to or use of the Services by or for the benefit of any unauthorized school, library, organization, or user; > Publish, broadcast, sell, use or provide access to the Service or any materials retrieved from the Service in any manner that will infringe the copyright or other proprietary rights of ProQuest or its licensors; > Download all or parts of the Service in a systematic or regular manner or so as to create a collection of materials comprising all or a material subset of the Service, in any form. > Store any information on the Service that violates applicable law or the rights of any third party.
- paulmd 3y agoGrandparent is right that people keep conflating copyright and licensing. That term doesn’t allow you to violate the copyright of a licensor, but, it also doesn’t rule out (eg) fair use. Fair use would have to be blocked via a license, and it’s going to be difficult to argue that someone agrees to a license merely by turning on their radio. Responding to unauthenticated internet requests with content is the internet equivalent of broadcast and similarly the LinkedIn case held that this did not allow LinkedIn to impose terms of service in a contract of adhesion in this fashion.
- bryanrasmussen 3y agoactually if you had a subscription you might be more screwed than if you crawled it with free access, since having the subscription means agreement to the TOS.
- bandrami 3y agoViolating their terms of service doesn't really matter in terms of copyright law though. The statistical properties of that text (which is what the engine snarfs in) aren't protected by copyright.