Processor list

This page contains detailed information on all processors offered by Document AI. You can see a list of all processors by solution type.

All Document AI processors adhere to the Data Processing and Security Terms.

Refer to the Managing processor versions documentation for more details. Also, specific processor limits apply in addition to overall product quotas and limits.

Digitize text

Enterprise Document OCR (Optical Character Recognition)

Description

Identify and extract text in different types of documents.

This processor allows you to identify and extract text, including handwritten text, from documents in more than 200 languages. The processor also uses machine learning to perform a quality assessment of a document based on the readability of its content.

Category Digitize
Functions OCR, Quality Analysis
Release stage General availability
Access status Public
Type in API OCR_PROCESSOR
Supported languages
Full list of languages
Language Name BCP 47 Tag Script Handwriting supported
Afrikaans af Latn
Albanian sq Latn
Arabic ar Arab
Armenian hy Armn
Belarusian be Cyrl
Bangla bn Beng
Bengali bn Beng
Bulgarian bg Cyrl
Catalan ca Latn
Chinese zh Hani
Croatian hr Latn
Czech cs Latn
Danish da Latn
Dutch nl Latn
English en Latn
Estonian et Latn
Filipino fil Latn
Finnish fi Latn
French fr Latn
German de Latn
Greek el Grek
Gujarati gu Gujr
Hebrew iw Hebr
Hindi hi Deva
Hungarian hu Latn
Icelandic is Latn
Indonesian id Latn
Italian it Latn
Japanese ja Jpan
Kannada kn Knda
Khmer km Khmr
Korean ko Kore
Lao lo Laoo
Latvian lv Latn
Lithuanian lt Latn
Macedonian mk Cyrl
Malay ms Latn
Malayalam ml Mlym
Marathi mr Deva
Nepali ne Deva
Norwegian no Latn
Persian fa Arab
Polish pl Latn
Portuguese (Portugal & Brazil) pt Latn
Punjabi pa Guru
Romanian ro Latn
Russian ru Cyrl
Serbian sr Cyrl
Slovak sk Latn
Slovenian sl Latn
Spanish es Latn
Swedish sv Latn
Tagalog tl Latn
Tamil ta Taml
Telugu te Telu
Thai th Thai
Turkish tr Latn
Ukrainian uk Cyrl
Vietnamese vi Latn
Yiddish yi Hebr
Processor versions
Version ID Release Channel Release Maturity Description
pretrained-ocr-v1.2-2022-11-10 Stable GA Frozen model version of v1.0: Model files, configurations, and binaries of a version snapshot frozen in a container image for up to 18 months.
pretrained-ocr-v2.0-2023-06-02 Stable GA Production-ready model specialized for document use cases. Includes access to all OCR add-ons.
pretrained-ocr-v2.1-2024-08-07 Stable GA The main areas of improvement for v2.1 are: better printed text recognition, more precise checkbox detection and more accurate reading order.
pretrained-ocr-v2.1.1-2025-01-31 Release candidate Public Preview v2.1.1 is similar to V2.1, and is available in all regions except: US, EU, and asia-southeast1.

For more information, see Managing processor versions.

Quotas and limits
Maximum pages (online/synchronous requests): 15
Maximum pages (batch/offline/asynchronous requests): 500
Maximum pages (imageless mode online/synchronous requests): 30
Uptraining
Sample Input File Open in new window.
Sample Output Open in new window.
Supported regions
  • asia-south1
  • asia-southeast1
  • australia-southeast1
  • eu
  • europe-west2
  • europe-west3
  • northamerica-northeast1
  • us
More information Enterprise Document OCR

Extract entities from documents

Refer to Sample datasets for sample labeled and unlabeled datasets to use for training.

Custom Extractor

Description

Extract fields from documents using generative AI or custom models; fine-tune models to accurately extract data from your documents.

Category Extract
Functions OCR, Entity Extraction
Release stage General availability
Access status Public
Type in API CUSTOM_EXTRACTION_PROCESSOR
Notes
  • If using generative AI for extraction, then:

    • Only the English language is officially supported.
    • Region availability is in the US, EU, northamerica-northeast1 and asia-southeast1.

Supported languages
Full list of languages
Language Name BCP 47 Tag Script Handwriting supported
Afrikaans af Latn
Arabic ar Arab
Azerbaijani az Latn
Azerbaijani (Cyrillic) az-Cyrl Cyrl
Belarusian be Cyrl
Bulgarian bg Cyrl
Bosnian bs Latn
Catalan ca Latn
Cebuano ceb Latn
Czech cs Latn
Welsh cy Latn
Danish da Latn
German de Latn
Greek el Grek
English en Latn
Esperanto eo Latn
Spanish es Latn
Estonian et Latn
Basque eu Latn
Persian fa Arab
Finnish fi Latn
Filipino fil Latn
French fr Latn
Irish ga Latn
Galician gl Latn
Hindi hi Deva
Croatian hr Latn
Haitian Creole ht Latn
Hungarian hu Latn
Indonesian id Latn
Icelandic is Latn
Italian it Latn
Hebrew iw Hebr
Japanese ja Jpan
Javanese jv Latn
Kazakh kk Cyrl
Korean ko Kore
Kyrgyz ky Cyrl
Latin la Latn
Lithuanian lt Latn
Latvian lv Latn
Macedonian mk Cyrl
Mongolian mn Cyrl
Marathi mr Deva
Malay ms Latn
Maltese mt Latn
Nepali ne Deva
Dutch nl Latn
Norwegian no Latn
Polish pl Latn
Pashto ps Arab
Portuguese (Portugal & Brazil) pt Latn
Romanian ro Latn
Russian ru Cyrl
Russian (Petrine Orthography) ru-PETR1708 Cyrl
Sanskrit sa Deva
Slovak sk Latn
Slovenian sl Latn
Albanian sq Latn
Serbian sr Cyrl
Swedish sv Latn
Swahili sw Latn
Tagalog tl Latn
Turkish tr Latn
Ukrainian uk Cyrl
Urdu ur Arab
Uzbek uz Latn
Uzbek (Cyrillic) uz-Cyrl Cyrl
Vietnamese vi Latn
Yiddish yi Hebr
Chinese simplified zh-Hans Hani
Chinese traditional zh-Hant Hani
Zulu zu Latn
Processor versions
Version ID Release Channel Release Maturity Description
pretrained-foundation-model-v1.5-2025-05-05 Stable GA Production-ready candidate powered by Gemini 2.5 Flash LLM. Recommended for those who want to experiment with newer models.
pretrained-foundation-model-v1.5-pro-2025-06-20 Stable GA Production-ready model powered by the Gemini 2.5 Pro LLM. Supports a quota of up to 30 pages per minute for online process requests. This model has improved quality compared to v1.5, and may have a higher latency.
pretrained-foundation-model-v1.5.1-2025-08-07 Release candidate Public Preview Public preview model powered by the Gemini 2.5 Flash LLM. This model has the same features as v1.5, and has improved adaptive few-shot learning.
pretrained-foundation-model-v1.6-pro-2025-12-01 Release candidate Public Preview Preview model powered by the Gemini 3 Pro LLM.
pretrained-foundation-model-v1.6-2026-01-13 Release candidate Public Preview Preview model powered by the Gemini 3 Flash LLM.
pretrained-foundation-model-v3.5-2026-05-26 Release candidate Public Preview Preview model powered by the Gemini 3.5 Flash LLM.

For more information, see Managing processor versions.

Quotas and limits
Maximum pages (online/synchronous requests): 15
Maximum pages (batch/offline/asynchronous requests): 200
Maximum pages (imageless mode online/synchronous requests): 30
Normalized data types

You can find more information in the Enrichment & normalization, and Create dataset pages.

Full list of normalized data types
  • dateTime as STRING
  • currency as STRING
  • money as google.type.Money
  • number as FLOAT or INTEGER
Uptraining
Sample Input File Open in new window.
Sample Output Open in new window.
Supported regions
  • asia-south1
  • asia-southeast1
  • australia-southeast1
  • eu
  • europe-west2
  • europe-west3
  • northamerica-northeast1
  • us
More information Custom Extractor

Form Parser

Description

Extract general key-value pairs (entity and checkbox), tables, and generic entities from documents in addition to OCR text.

This processor applies advanced machine learning technologies to extract key-value pairs, checkboxes, and tables from documents more than 200 languages. This processor also leverages deep learning models to extract 11 generic entities that are common in various document types.

Category Extract
Functions OCR, Form Parsing, Entity Extraction
Release stage General availability
Access status Public
Type in API FORM_PARSER_PROCESSOR
Supported languages
Full list of languages
Language Name BCP 47 Tag Script Handwriting supported
Afrikaans af Latn
Albanian sq Latn
Arabic ar Arab
Belarusian be Cyrl
Catalan ca Latn
Chinese zh Hani
Croatian hr Latn
Czech cs Latn
Danish da Latn
Dutch nl Latn
English en Latn
Estonian et Latn
Filipino fil Latn
Finnish fi Latn
French fr Latn
German de Latn
Hebrew iw Hebr
Hindi hi Deva
Hungarian hu Latn
Icelandic is Latn
Indonesian id Latn
Italian it Latn
Japanese ja Jpan
Korean ko Kore
Latvian lv Latn
Lithuanian lt Latn
Macedonian mk Cyrl
Malay ms Latn
Marathi mr Deva
Nepali ne Deva
Norwegian no Latn
Persian fa Arab
Polish pl Latn
Portuguese (Portugal & Brazil) pt Latn
Romanian ro Latn
Russian ru Cyrl
Serbian sr Cyrl
Slovak sk