UiPath Documentation
document-understanding
2023.4
false
Document Understanding User Guide

OCR engines

An OCR Engine is used in the Digitization component, to identify text in a file, when native content is not available.

Note:

The images that need to be processed should have a resolution range of:

  • min: 50 x 50 pixels
  • max: 9000 x 9000 pixels

Here is a selection of OCR Engines that you can choose from, according to your needs, throughout the Document UnderstandingTM Framework.

OCR Engine Activity Pack Debug Logs Format in Logs Folder Reports Confidence
UiPath Extended Languages OCRUiPath.OCR.Activities${date:format=yyyy-MM-dd}

Yes

UiPath Document OCR

UiPath.OCR.Activities

${date:format=yyyy-MM-dd}

Yes

OCR for Chinese, Japanese and KoreanUiPath.Core.Activities.CjkOCR${date:format=yyyy-MM-dd}

Yes

OmniPage OCR

UiPath.OmniPage.Activities

${date:format=yyyy-MM-dd}

Yes

Google Cloud Vision OCR

UiPath.UIAutomation.Activities

${date:format=yyyy-MM-dd}

No if DetectionMode is set to TextDetection (default)

Yes if DetectionMode is set to DocumentTextDetection

Microsoft Azure Computer Vision OCR

UiPath.UIAutomation.Activities

${date:format=yyyy-MM-dd}

No if UseReadAPI is not selected (default)

Yes if UseReadAPI is selected

Microsoft OCR

UiPath.UIAutomation.Activities

${date:format=yyyy-MM-dd}

No

Tesseract OCR

UiPath.UIAutomation.Activities

${date:format=yyyy-MM-dd}

Yes

Note:

When debugging errors, you can always visit the logs folder and check the relevant OCR log files. Read more about logging here.

Was this page helpful?

Connect

Need help? Support

Want to learn? UiPath Academy

Have questions? UiPath Forum

Stay updated