50
Chapter 4
Languages
The program can read over 110 languages with three alphabets: Latin,
Greek and Cyrillic. See the list in the OCR panel of the Options dialog
box. It shows which languages have dictionary support. A listing is also
provided on the ScanSoft web site.
In addition to user dictionaries, specialized dictionaries are available for
certain professions (currently medical, legal and financial) for some
languages. See the list and make selections in the OCR panel of the
Options dialog box.
Training
Training is the process of changing the OCR solutions assigned to
character shapes in the image. It is useful for uniformly degraded
documents or when an unusual typeface is used throughout a document.
OmniPage 15 offers two types of training: manual training and automatic
training (IntelliTrain). Data coming from both types of training are
combined and available for saving to a training file.
When you leave a page on which training data was generated, you will be
asked how to apply it to other existing pages in the document.
Manual training
To do manual training, place the insertion point in front of the character
you want to train, or select a group of characters (up to one word) and
choose Train Character... from the Tools menu or the shortcut menu. You
will see an enlarged view of the character(s) to be trained, along with the
current OCR solution. Change this to the desired solution and click OK.
The program takes this training and examines the rest of the page. If it
finds candidate words to change, the Check Training dialog box lists
these. Incorrect words should be re-trained before the list is approved.
Summary of Contents for OMNIPAGE 15
Page 1: ......