4 ms·
Publicly available web pages can always be legally indexable and searchable, right? An index is a derivative-enough work that copyright doesn't apply.
by Khelavaster 2y ago
Publicly available web pages can always be legally indexable and searchable, right?
An index is a derivative-enough work that copyright doesn't apply.
- petterroea 2y agoI agree with you, but I feel you are implying that we should not respect robots.txt any more: https://www.reddit.com/robots.txt https://www.reddit.com/robots.txt The idea of ignoring robots.txt when it is abused for greed sounds okay, but it would be much better if the industry was in a spot where we respected each others' requests, and peers don't make requests that put legitimate, honest actors at a disadvantage compared to rouge crawlers
- zarzavat 2y agoIn the US the courts have ruled on this and their answer was “eh?”: https://en.m.wikipedia.org/wiki/HiQ_Labs_v._LinkedIn https://en.m.wikipedia.org/wiki/HiQ_Labs_v._LinkedIn There are 195 other countries you could scrape from though. As for whether an index constitutes copyright infringement, that depends on its construction, but in the US generally not.
- jerrygoyal 2y agoso Bing can create a different entity in India for crawling and use that to show reddit content in search results? I'm sure they would have already done it it if it was a possibility.
- zarzavat 2y agoIf they bought the index from a company that had compiled it legally according to the law of another country they might be okay. The terms of service bind the legal entity that is accessing Reddit’s servers, whereas reproduction of the data is governed by copyright law. Microsoft would still have to abide by US copyright law but there’s some wiggle room in terms of who is doing the scraping.