4 ms·
Would love to see what clusters from PubMed would look like. Anyone planning to run this on it?
by JunkDNA 14y ago
Would love to see what clusters from PubMed would look like. Anyone planning to run this on it?
- skram 14y agoAgreed. Would love to see this too. IIRC they have a pretty easy to use API but as far as data dumps (according to a quick search and http://www.nlm.nih.gov/bsd/sample_records_avail.html http://www.nlm.nih.gov/bsd/sample_records_avail.html) it appears they only provide XML whereas the current code requires XML.
- JunkDNA 14y agoWhile it's true you'd get XML from NIH, as a first pass, just extracting the MEDLINE titles and abstracts into plaintext and then running this over them would be enough. You need to complete a license agreement (annoyingly) to download them. But there's no fee or anything.
- dbaupp 14y ago> it appears they only provide XML whereas the current code requires XML Is there a typo here, or am I just reading this wrong?
- skram 14y agotypo -- second instance of "XML" should be ".txt"
- deleted 14y ago[deleted]
- omershapira 14y agoAnyone who wants to do that can contact me (contact form and email can be found there) and I'll happily publish the results.