OptionaldocumentDocument-level metadata extracted from the source container (best-effort). Present only when the extractor was able to read it.
OptionallongPath to the original file in cloud storage, if applicable.
OptionalocrPages the OCR provider reported as processed (billed), when it reports one. Can exceed
pages.length — a blank page is billed but returns no content.
Array of pages containing extracted data.
OptionalspreadsheetStructural information about the sheets when the document is a parseable spreadsheet (best-effort). Presence marks the document for the spreadsheet summary-embedding lane.
The full result of extracting data from a document, including all pages and any extracted images.