Home · Glossary · OCR (Optical Character Recognition)
Enterprise AI glossary · Document Intelligence (IDP/OCR)

OCR (Optical Character Recognition)

The classical step of converting an image of text into machine-readable characters — the foundation layer underneath any document processing pipeline.

Definition

What OCR (Optical Character Recognition) means in practice

OCR is the step of converting an image (or PDF page rendered as an image) into machine-readable text. Modern OCR — Tesseract, PaddleOCR, AWS Textract, Azure Form Recognizer, Google Document AI — handles printed text reliably and handwritten text passably. The hard problems are downstream: text alone is not structured data. "4,28,940" on a row labelled "AMOUNT" needs to become an integer field tied to an invoice record. That work is IDP, not OCR. Treating OCR as the destination rather than the foundation is the most common reason enterprise document-automation programmes stall after the first wave of straight-through documents.

Go deeper
Document intelligence (IDP) →

All 62 terms, in plain language

Sovereign AI, RAG, agentic AI, IDP, MLOps and the regulations that shape enterprise AI.

Browse the glossary →Talk to an engineer →