AI Fundamentals

OCR-VLM

A vision-language model specialized for optical character recognition — reading text, tables, forms, and structured documents from images. Modern VLMs handle most OCR as a subtask, replacing dedicated OCR pipelines for enterprise document workflows.

Related terms

Next steps

Where this fits

OCR-VLM is part of the AI Fundamentals vocabulary used in the Generative AI Maturity Framework. See the full glossary for the complete set of 149 defined terms, or take the free maturity assessment to see where your organisation stands.