2 ms·
How did they get the plain text as metadata? Was the scanning equipment doing OCR and setting that?
by brc 11y ago
How did they get the plain text as metadata? Was the scanning equipment doing OCR and setting that?
- danso 11y agoI believe the U.S. agencies use ABBYY FineReader, which does a pretty good job with OCR and text resolution. The U.S. Senate used it when releasing the CIA torture docs awhile ago: http://www.nytimes.com/interactive/2014/12/09/world/cia-torture-report-document.html?_r=0 http://www.nytimes.com/interactive/2014/12/09/world/cia-tort...