UiPath Documentation
document-understanding
2022.4
true
Document Understanding 用户指南
重要 :
新发布内容的本地化可能需要 1-2 周的时间才能完成。

部署 UiPath 文档 OCR

在 AI Center 中创建 UiPath 文档 OCR ML 包。

对于在线安装,“开箱即用包”部分已包含“UiPath 文档 OCR”模型。转到“ML 包”>“开箱即用包”>“UiPath Document Understanding”>“UiPath 文档 OCR”,然后单击“提交”。

For offline installation, go to the ML Packages tab from the left sidebar of AI Center and create a new package. Name the package and upload the package that you have downloaded from this page. Choose JSON input type, and the corresponding Python language. Create package.

备注:

When creating a UiPathDocumentOCR ML Package in AI Center, it cannot be named ocr or OCR. Make sure to choose another name.

转到“ML 技能”,并为您创建的 UiPath 文档 OCR 包创建新的 ML 技能。

请使用高级基础架构设置以更新部署来更新副本(理想情况下,副本数应等于节点数),并最大化 CPU(至少 4 个)和 RAM 请求(如果您未使用 GPU 计算机,或 UiPath 文档 OCR 处理速度较慢,并且可能会失败)。

OCR 引擎需要 GPU 上才能实现最佳性能,建议用于生产工作负载。但是,如果无可用 GPU,它仍可以在 CPU 上运行,但需要比默认设置更多的资源。高级基础架构设置应进行如下调整:

  • Replicas: increase if there is concurrent usage of UiPathDocumentOCR. If you are using UiPathDocumentOCR to do imports on a single Data Labeling session at a time and the UiPathDocumentOCR is not used in other UiPath workflows then 1 replica suffices. Otherwise, the number of replicas needs to be increased. There is no "magic" number here, you need some trial and error. Do not use more than 2 replicas on a single node installation. Ideally, replica count should equal the number of nodes in the cluster (1 replica/node). If more parallelism is needed, increasing the CPU helps

  • CPU:至少应为 4 个(对于每个副本)。请确保您拥有适当的资源。没有一个“确定无误”的数字,但更多的 CPU 意味着更快的处理时间。您需要在特定场景下测试是否足够。

ML 技能可能需要长达 30 分钟才能准备就绪。您可能需要刷新 AI Center 页面才能查看状态更改。ML 技能可用后,双击 ML 技能并转到“修改当前部署”。

打开开关,将 ML 技能设为公开。您可能需要等待几分钟才能刷新页面。

双击 ML 技能并复制 URL,即 UiPath 文档 OCR 的端点,以供以后使用。

恭喜!您已在 AI Center 上成功部署 UiPath 文档 OCR

此页面有帮助吗?

连接

需要帮助? 支持

想要了解详细内容? UiPath Academy

有问题? UiPath 论坛

保持更新