Re: Cost of no OCR for extended Latin
| Newsgroups | gmane.text.unicode.devel |
|---|---|
| Message-ID | <OFFEFF04B8.C4E65A17-ON8625737F.0052EFD2-8625737F.00534040@notes.sil.org> |
> David Starner wrote on 10/25/2007 05:41:19 AM: > > On 10/25/07, Don Osborn <[email protected]> wrote: > > Is anyone aware of an OCR system that recognizes extended Latin characters > > from say Extended A&B, IPA, and Extended Additional ranges? That is for any > > language (orthography) including these characters? > > ABBYY offers most of Extended A and some of Extended B and Additional. > The list of supported languages is > <http://www.abbyy.com/finereader8/?param=44927>, which should map to > the list of supported characters. It would be hard to impossible to > create and test an OCR without a substantial corpus of material using > a character; I suspect many languages are on ABBYY's list only because > the orthography is a subset of those supported for other reasons. Quoting two different colleagues of mine: "I recommend FineReader (www.finereader.com) from Abbyy Software. While OmniPage is good, FineReader is better--the best OCR software at an affordable price...FineReader can handle special characters better than other OCR programs." and "I heartily recommend FineReader. It can be "trained" to recognize speciality characters, and it is surprisingly accurate - about 99% - which means that 1% of the document will require manual corrections." Lorna