Question

3
Replies
203
Views
JayanthiK4197 Member since 2018 8 posts
Bank Of America
Posted: 11 months ago
Last activity: 11 months 2 weeks ago

Can Document OCR not get text out of a scanned pdf?

I have tried processToXml and ProcessToPdf and tried putting ProcessToPdf before each of these and thried everything with and without ocrImagesAndText being true. everything just returns false. I am trying to get text out of a pdf produced by scanning a paper document, but there are even some pdfs the regular pdf connector can read that document ocr cannot, unless I just cannot sort out how to use it. I can make it get text from images in word documents and it can get text out of a pdf I make by doing a print to pdf, so I know I am not doing everything wrong. can this component actually not get text from a scanned pdf?

Robotic Process Automation
Share this page LinkedIn