Back to the catalogue
IngestionParsing

Document OCR

Turns scanned pages and photographed documents into searchable text, keeping the reading order.

Half of the documents worth indexing arrived as a scan: a signed contract, a policy printed in 2019, a supplier invoice photographed on a phone.

OCR runs as part of ingestion, so the text lands in the same knowledge base as everything else and answers cite the page it came from. It is also available to an agent as a step, for the times a document turns up mid-run.

Use Document OCR in your own workspace.

Create a workspace, connect a source, and put this to work in a few minutes.

Book a demo