3 ms·
Not that hard? You'd need to process each page then data mine it.
by dontbeunethical 7y ago
Not that hard?
You'd need to process each page then data mine it.
- eof 7y agoA company that can build a game engine should probably be able to crawl a site and save an html dump?
- biggestdecision 7y agoJust archive.org every page on the wiki. Their api will let you do 15/minute.
- smacktoward 7y agoHTTrack (https://www.httrack.com/ https://www.httrack.com/) makes tasks like this trivial.
- imtringued 7y agoIt doesn't run a full browser engine so it won't work with the vast majority of websites.