how to extract text reliably from a PDF file?
Yaga <[email protected]> Wed, 21 Aug 2002 15:58:41 +0000 (UTC)
| Newsgroups | gmane.comp.printing.ghostscript.bugs |
|---|---|
| Organization | Yaga Soop |
| Message-ID | <[email protected]> |
Hi list, I want to use GhostScript GNU, to reliably extract text from a PDF file, into a plain text file. The text file could contain Unicode or ASCII, depending on the PDF file. GhostScript seems very complicated. What is the technique?