Question

3
Replies
214
Views
JayanthiK4197 Member since 2018 8 posts
Bank Of America
Posted: 1 year ago
Last activity: 1 year 1 month ago
Closed

Can Document OCR not get text out of a scanned pdf?

I have tried processToXml and ProcessToPdf and tried putting ProcessToPdf before each of these and thried everything with and without ocrImagesAndText being true. everything just returns false. I am trying to get text out of a pdf produced by scanning a paper document, but there are even some pdfs the regular pdf connector can read that document ocr cannot, unless I just cannot sort out how to use it. I can make it get text from images in word documents and it can get text out of a pdf I make by doing a print to pdf, so I know I am not doing everything wrong. can this component actually not get text from a scanned pdf?

Robotic Process Automation
Moderation Team has archived post
Share this page LinkedIn