Google Gemini Layout Parser 1.6
Google Cloud's Gemini-assisted parser for preserving document hierarchy and creating context-rich chunks for enterprise search and RAG.
Google Cloud's Gemini-assisted parser for preserving document hierarchy and creating context-rich chunks for enterprise search and RAG.
Plain-English overview
Gemini Layout Parser 1.6 combines Google OCR with Gemini 3.0 Flash. It identifies headings, paragraphs, lists, tables, figures, headers, and footers, then builds a document tree so a retrieved passage can keep the context of the section it came from.
This is a parsing and chunking service rather than a general chat model. It is especially useful when flat OCR would separate a table from its headers or a paragraph from its parent section.
Category comparison
These are provider-published specifications, not Cody benchmark scores. Follow the linked sources for current limits and endpoint-specific exceptions.
Pricing & comparisons
Set your usage. Your estimate updates as you type.
Counts pages, not files. A 10-page PDF counts as 10 pages. The estimate uses the document-processing rate available for each model.
Google Gemini Layout Parser 1.6
Google Cloud
Estimated total (USD)
For the usage above · USD · API pricing, not a subscription
Estimates exclude taxes, tools, cache storage/writes, free allowances and custom discounts. Image estimates cover output only, not prompt or reference-image charges. Quality modes differ by model. Unlisted settings are not treated as free.
API and provider access
Availability
Release-candidate access through Google Cloud Document AI.
This version uses the global Vertex AI Gemini endpoint and is explicitly not compliant with Google's data-residency standard, even when called through US or EU endpoints.
Check live availabilityData and training
Google says Document AI customer documents and predictions are not used to train its models. Online documents are processed in memory without disk persistence; batch inputs have a failsafe retention of up to one day.
This is a concise reading of the cited provider material, not legal advice. A third-party gateway can have different storage, routing, training, and residency terms from the model maker's direct API.
Read the provider policyFrequently asked questions
Google Cloud's Gemini-assisted parser for preserving document hierarchy and creating context-rich chunks for enterprise search and RAG. Gemini Layout Parser 1.6 combines Google OCR with Gemini 3.0 Flash. It identifies headings, paragraphs, lists, tables, figures, headers, and footers, then builds a document tree so a retrieved passage can keep the context of the section it came from.
Google Gemini Layout Parser 1.6 was released on January 13, 2026 according to the cited provider materials.
Release-candidate access through Google Cloud Document AI. The access routes listed in this guide are Google Cloud.
$10 / 1,000 pages. Google Cloud's Layout Parser price includes initial chunking. Re-chunking is billed separately.
Release-candidate access through Google Cloud Document AI. This version uses the global Vertex AI Gemini endpoint and is explicitly not compliant with Google's data-residency standard, even when called through US or EU endpoints.
Google says Document AI customer documents and predictions are not used to train its models. Online documents are processed in memory without disk persistence; batch inputs have a failsafe retention of up to one day. The policy belongs to the provider route and account terms, so verify it again before production use.
Related comparisons
Cohere's compact document parser for turning enterprise PDFs, presentations, and scans into retrieval-ready Markdown and layout metadata.
Open full comparisonMistral's current Document AI OCR model for preserving reading order, tables, structure, locations, and confidence across multilingual files.
Open full comparisonGoogle Cloud's Gemini-powered Document AI processor for extracting the exact fields and derived values defined in a business schema.
Open full comparison