3 ms·
That seems good -- but I can imagine Google preferring to crawl the page, rather than receive it by API (so it's more likely to be what the user's going to see)
by aamar 16y ago
That seems good -- but I can imagine Google preferring to crawl the page, rather than receive it by API (so it's more likely to be what the user's going to see).
I think Google can handle the scaling problem; one not-great solution: ignore notifications except for those from people who are being scraped and need it.
Still, it's kind of a shame that webmasters have to worry about any of this.
- mootothemax 16y agoThat seems good -- but I can imagine Google preferring to crawl the page, rather than receive it by API I'd agree with you, but can think of edge cases where naughty sites A, B and C submit new URLs within an arbitrary amount of time of the new content being published. In that case there'd be no way for Google to tell who published the content first other than to have a big list of original content publishers - and I think that list'd get messy fast.
- aamar 16y agoYou're right, there's an exploit there for the re-publishers. Great point.