4 ms·
Isn't web crawling essentially "training data" ?
by AK42 3y ago
Isn't web crawling essentially "training data" ?
- jonny_eh 3y agoJust because something is published on the web, it doesn't mean it's free of copyright protections.
- TerrifiedMouse 3y agoNope. That just indexing, i.e. making an index. That said, Google does cache other people's content and walks a fine line doing so - there is also that AMP thing. They probably haven't gotten sued because they don't affect those websites' ad revenue - don't know how that works.
- sebzim4500 3y agoAMP is opt-in, isn't it? Presumably the website has agreed to their data being distributed in that way.