2 ms·
I can't believe wayback machine has this. I ran a text-only scraper for ~200 websites back in the early 2000s. https://web.archive.org/web/20060803053549/http
by mrmekon 9y ago
I can't believe wayback machine has this. I ran a text-only scraper for ~200 websites back in the early 2000s.
https://web.archive.org/web/20060803053549/http://inr.cjb.net:80/ https://web.archive.org/web/20060803053549/http://inr.cjb.ne...
Most were RSS feeds, since those used to exist. The rest were hand-written parsers, which were highly temperamental.