Enterprise Document OCR (Optical Character Recognition)
Description
Identify and extract text in different types of documents.
This processor allows you to identify and extract text, including handwritten text, from documents in more than 200 languages. The processor also uses machine learning to perform a quality assessment of a document based on the readability of its content.
Only the English language is officially supported.
Region availability is in the US, EU, northamerica-northeast1 and asia-southeast1.
Supported languages
Full list of languages
Language Name
BCP 47 Tag
Script
Handwriting supported
Afrikaans
af
Latn
Arabic
ar
Arab
Azerbaijani
az
Latn
Azerbaijani (Cyrillic)
az-Cyrl
Cyrl
Belarusian
be
Cyrl
Bulgarian
bg
Cyrl
Bosnian
bs
Latn
Catalan
ca
Latn
Cebuano
ceb
Latn
Czech
cs
Latn
Welsh
cy
Latn
Danish
da
Latn
German
de
Latn
Greek
el
Grek
English
en
Latn
Esperanto
eo
Latn
Spanish
es
Latn
Estonian
et
Latn
Basque
eu
Latn
Persian
fa
Arab
Finnish
fi
Latn
Filipino
fil
Latn
French
fr
Latn
Irish
ga
Latn
Galician
gl
Latn
Hindi
hi
Deva
Croatian
hr
Latn
Haitian Creole
ht
Latn
Hungarian
hu
Latn
Indonesian
id
Latn
Icelandic
is
Latn
Italian
it
Latn
Hebrew
iw
Hebr
Japanese
ja
Jpan
Javanese
jv
Latn
Kazakh
kk
Cyrl
Korean
ko
Kore
Kyrgyz
ky
Cyrl
Latin
la
Latn
Lithuanian
lt
Latn
Latvian
lv
Latn
Macedonian
mk
Cyrl
Mongolian
mn
Cyrl
Marathi
mr
Deva
Malay
ms
Latn
Maltese
mt
Latn
Nepali
ne
Deva
Dutch
nl
Latn
Norwegian
no
Latn
Polish
pl
Latn
Pashto
ps
Arab
Portuguese (Portugal & Brazil)
pt
Latn
Romanian
ro
Latn
Russian
ru
Cyrl
Russian (Petrine Orthography)
ru-PETR1708
Cyrl
Sanskrit
sa
Deva
Slovak
sk
Latn
Slovenian
sl
Latn
Albanian
sq
Latn
Serbian
sr
Cyrl
Swedish
sv
Latn
Swahili
sw
Latn
Tagalog
tl
Latn
Turkish
tr
Latn
Ukrainian
uk
Cyrl
Urdu
ur
Arab
Uzbek
uz
Latn
Uzbek (Cyrillic)
uz-Cyrl
Cyrl
Vietnamese
vi
Latn
Yiddish
yi
Hebr
Chinese simplified
zh-Hans
Hani
Chinese traditional
zh-Hant
Hani
Zulu
zu
Latn
Processor versions
Version ID
Release Channel
Release Maturity
Description
pretrained-foundation-model-v1.5-2025-05-05
Stable
GA
Production-ready candidate powered by Gemini 2.5 Flash LLM. Recommended for those who want to experiment with newer models.
pretrained-foundation-model-v1.5-pro-2025-06-20
Stable
GA
Production-ready model powered by the Gemini 2.5 Pro LLM. Supports a quota of up to 30 pages per minute for online process requests. This model has improved quality compared to v1.5, and may have a higher latency.
pretrained-foundation-model-v1.5.1-2025-08-07
Release candidate
Public Preview
Public preview model powered by the Gemini 2.5 Flash LLM. This model has the same features as v1.5, and has improved adaptive few-shot learning.
pretrained-foundation-model-v1.6-pro-2025-12-01
Release candidate
Public Preview
Preview model powered by the Gemini 3 Pro LLM.
pretrained-foundation-model-v1.6-2026-01-13
Release candidate
Public Preview
Preview model powered by the Gemini 3 Flash LLM.
pretrained-foundation-model-v3.5-2026-05-26
Release candidate
Public Preview
Preview model powered by the Gemini 3.5 Flash LLM.
Extract general key-value pairs (entity and checkbox), tables, and generic entities from documents in addition to OCR text.
This processor applies advanced machine learning technologies to extract key-value pairs, checkboxes, and tables from documents more than 200 languages. This processor also leverages deep learning models to extract 11 generic entities that are common in various document types.