16.0
52 Chapter 4
Languages
The program can read over 110 languages with three alphabets:
Latin, Greek and Cyrillic. See the list in the OCR panel of the
Options dialog box. It shows which languages have dictionary
support. A listing is also provided on the Nuance web site.
In addition to user dictionaries, specialized dictionaries are
available for certain professions (currently medical, legal and
financial) for some languages. See the list and make selections in the
OCR panel of the Options dialog box.
Training
Training is the process of changing the OCR solutions assigned to
character shapes in the image. It is useful for uniformly degraded
documents or when an unusual typeface is used throughout a
document. OmniPage 16 offers two types of training: manual
training and automatic training (IntelliTrain). Data coming from
both types of training are combined and available for saving to a
training file.
When you leave a page on which training data was generated, you
will be asked how to apply it to other existing pages in the
document.
Manual training
To do manual training, place the insertion point in front of the
character you want to train, or select a group of characters (up to
one word) and choose Train Character... from the Tools menu or the
shortcut menu. You will see an enlarged view of the character(s) to
be trained, along with the current OCR solution. Change this to the
desired solution and click OK. The program takes this training and
examines the rest of the page. If it finds candidate words to change,