4 ms·
Because parsing the original metadata from Project Gutenberg is time consuming to write. I wasn't going to submit it to HN until I had an index/search api, but
by sethish 12y ago
Because parsing the original metadata from Project Gutenberg is time consuming to write. I wasn't going to submit it to HN until I had an index/search api, but someone beat me to it.
- arafalov 12y agoBut what/where is the metadata? Is it functionally equivalent to the Gutenberg's info (e.g. in the RDF dump). Or something else? I was looking to write an alternative search for Gutenberg, based on the RDF dump, so would be happy to collaborate/discuss ideas.
- sethish 12y agoYep. The RDF/XML data. I have a mirror of it on github: https://github.com/sethwoodworth/PG_rdf_metadata https://github.com/sethwoodworth/PG_rdf_metadata I would love to have a complete python parser for the metadata. I strongly recommend collaborating with the Gutenberg package posted to HN a few weeks ago (and his rdf branch): https://github.com/c-w/Gutenberg/tree/migrate-to-rdf https://github.com/c-w/Gutenberg/tree/migrate-to-rdf GITenberg has a mailing list and would love to have you! https://groups.google.com/forum/#!forum/gitenberg-project https://groups.google.com/forum/#!forum/gitenberg-project