4 ms·
Some of my primary interests are building websites that can analyse other websites and for that, I need a keyword index and access to the underlying crawl data
by struct 11y ago
Some of my primary interests are building websites that can analyse other websites and for that, I need a keyword index and access to the underlying crawl data (as in, a response that can point me to the exact file offsets that contain the pages). Think [1] but with keywords instead of URLs.
[1] http://index.commoncrawl.org/ http://index.commoncrawl.org/
- sylvinus 11y agoThanks! We probably won't expose the underlying crawl data ourselves but being able to reference it just like [1] does is indeed important.