The Vision API can detect and transcribe text from PDF and TIFF files stored in Cloud Storage.
Document text detection from PDF and TIFF must be requested using the
files:asyncBatchAnnotate function, which performs an offline (asynchronous)
request and provides its status using the operations resources.
Output from a PDF/TIFF request is written to a JSON file created in the specified Cloud Storage bucket.
Limitations
The Vision API accepts PDF/TIFF files up to 2000 pages. Larger files will return an error.
Authentication
API keys are not supported for files:asyncBatchAnnotate requests. See
Using a service account for
instructions on authenticating with a service account.
The account used for authentication must have access to the Cloud Storage
bucket that you specify for the output (roles/editor or
roles/storage.objectCreator or above).
You can use an API key to query the status of the operation; see Using an API key for instructions.
Document text detection requests
Currently PDF/TIFF document detection is only available for files stored in Cloud Storage buckets. Response JSON files are similarly saved to a Cloud Storage bucket.
gs://cloud-samples-data/vision/pdf_tiff/census2010.pdf,