extract document text.md

Extract Document Text

Experimental Feature

This feature is Experimental and may change based on user feedback and testing. Share your thoughts via our chatbot to help us improve it.

The Extract Document Text block extracts text from supported documents and images using OCR (Optical Character Recognition). This is useful when you need to validate or reuse text from scanned PDFs or image files inside an automation flow (for example, extracting text from a scanned contract before checking an approval keyword). This feature is available starting from Release 2025.1.486.

Note:

When fully expanded, the Extract Document Text block displays the following properties:

Quick-start

  1. Drag Extract Document Text onto the canvas.

  2. Connect the block in the flow and specify Source type , then provide the file in File Input.

  3. Run the flow when it's ready.

Building block parameters

Parameters

Note: All cases have a global timeout configured in the Settings panel. This is unrelated to the timeout of a single building block. However, a running case will automatically be cancelled if it runs for longer than the global timeout.

Resources

Topic Description
Flows FAQ Common questions about creating, running, and managing flows in Leapwork.
Flows Troubleshooting Guidelines and solutions for identifying and fixing issues that occur when building or running flows in Leapwork.
++Customer Portal Add-ons++ Customer portal section to activate your cloud blocks.