5 ms·
My grandmother, who passed away last week, was an author. She self-published two novels through a printer who typeset her books and printed a few hundred copie
by caryme 15y ago
My grandmother, who passed away last week, was an author. She self-published two novels through a printer who typeset her books and printed a few hundred copies. I don't know what happened to the original text files she gave her printer and I know she never got a digital typeset copy.
I've been wanting to re-release her novels as ebooks, but haven't had a way to digitize them. This is perfect for me.
- caryme 15y agoBy the way, if anyone is interested, I'm also republishing my grandmother's short, inspirational writings on this blog: http://goldenwriter.caryme.com/ http://goldenwriter.caryme.com/.
- davidw 15y agoIdeally you'd get something besides a PDF, which is a pain in the neck to turn into a 'real' eBook.
- caryme 15y agoI'll probably OCR the PDF and see where I can go from there. It's a big jump start that I couldn't have easily done on my own.
- ScottBurson 15y agoThe 1DollarScan website says they do the OCR for you. They also claim that the result is searchable, which of course it couldn't be if they didn't OCR it.
- caryme 15y agoI may still want to OCR it myself for more control over the process (particularly in identifying and correcting OCR errors). In my experience, there's a difference in expectation of quality between OCR to make a PDF searchable and OCR to generate a standalone text file.
- CamperBob 15y agoWhy is that? .PDF isn't proprietary enough?
- burgerbrain 15y agoPDF is not proprietary.
- davidw 15y agoPDF is really a display format, whereas epub/mobi are more HTMLish in that they are not quite so specific in how they want things displayed, and thus can be flowed into different screen sizes pretty easily.
- jberryman 15y agoWhat a great thing to do.
- deleted 15y ago[deleted]
- jianshen 15y agoI was really hoping that the Google Books project would solve this problem for regular people. I know they take collections from public libraries and OCR them. Why not do the same for private collections?