This page documents production updates to Document AI. We recommend that Document AI developers periodically check this list for any new announcements.
You can see the latest product updates for all of Google Cloud on the Google Cloud page, browse and filter all release notes in the Google Cloud console, or programmatically access release notes in BigQuery.
To get the latest product updates delivered to you, add the URL of this page to your feed reader, or add the feed URL directly.
July 17, 2026
Custom extractor model
pretrained-foundation-model-v3.5-2026-05-26 powered by Gemini 3.5
Flash LLM is available in Preview.
This processor version has ML processing capabilities in the US and EU.
For more information about available models, see the custom extractor page.
June 04, 2026
Custom extractor offers document validation and correction in Preview.
This feature allows you to enhance extraction accuracy with validation rules and document data using Common Expression Language (CEL) dialect.
For more information, see CEL dialect for document validation.
May 27, 2026
Layout parser image and table annotations is in General Availability (GA).
Layout parser can identify if there are images or tables in parsed documents. When found, images and tables are annotated as a descriptive block of text with the information depicted in the image and table.
March 31, 2026
Upgrading fine tuned custom extractor processors is now available in Preview.
The feature allows you to fine tune a new processor version with a newer base version, while keeping the configurations of the previously fine-tuned processor version selected. This is available through the UI in the Deploy & use tab in the console.
This is currently supported for upgrading pretrained-foundation-model-v1.4-2025-02-05
to pretrained-foundation-model-v1.5-2025-05-05.
For more information, see training overview.
March 27, 2026
Custom splitter model
pretrained-splitter-v1.5-2025-07-14 is available in
General Availability (GA).
March 23, 2026
Custom classifier models
pretrained-classifier-v1.6-2026-03-09 and pretrained-classifier-v1.6-pro-2026-03-09
are available in Preview.
Custom splitter models
pretrained-splitter-v1.6-2026-03-09 and pretrained-splitter-v1.6-pro-2026-03-09
are available in Preview.
March 03, 2026
Custom classifier model pretrained-classifier-v1.5-2025-08-05
is available as General Availability (GA).
For more information about available models, see the custom classifier page.
February 17, 2026
Document AI legacy processors will be discontinued on June 30, 2026. To preempt the risk of service failure while using legacy processors, we recommend transitioning to more stable, higher-quality processors.
The affected versions are:
| Type | Version |
|---|---|
| Identity parsers | pretrained-us-passport-v1.0-2021-06-14pretrained-fr-driver-license-v1.0-2021-06-14 |
| Tax and finance parsers | pretrained-1099misc-v1.1-2021-12-10pretrained-1099nec-v1.0-2021-08-11pretrained-1099r-v2.0-2022-07-25pretrained-1099int-v1.1-2021-12-10pretrained-ssa1099-v1.0-2021-08-09pretrained-1099g-v1.0-2021-05-27pretrained-1099g-v1.1-2021-12-10pretrained-1120-v3.0-2022-04-26pretrained-w9-v1.0-2020-09-25pretrained-w9-v1.1-2021-12-10pretrained-w9-v1.2-2022-01-27pretrained-w9-v2.0-2022-06-23 |
| Mortage and banking parsers | pretrained-mortgage-statement-v1.0-2021-10-17
|
| Procurement | pretrained-utility-v1.1-2021-04-09pretrained-utility-v1.2-2022-12-15 |
| Splitting | pretrained-procurement-splitter-v1.1-2021-04-09pretrained-procurement-splitter-v1.2-2022-08-19pretrained-lending-document-split-v1.0-2021-12-08pretrained-lending-document-split-v2.0-2021-12-09 |
| Summary | pretrained-foundation-model-v1.0-2023-08-22 |
To ensure uninterrupted service and benefit from improved extraction quality, we recommend you migrate to the following later versions before June 30, 2026:
Enterprise Document OCR: Migrate to
pretrained-ocr-v2.1-2024-08-07.Expense parser: Migrate to
pretrained-expense-v1.3.2-2024-09-11.Custom classifier: Migrate to
pretrained-classifier-v1.5-2025-08-05.Custom splitter: Migrate to
pretrained-splitter-v1.5-2025-07-14.Invoice parser: Migrate to
pretrained-invoice-v2.0-2023-12-06.Pay slip parser: Migrate to
pretrained-paystub-v3.0-2023-12-06.Bank statement parser: Migrate to
pretrained-bankstatement-v5.0-2023-12-06.
To learn more about the migration process, refer to Manage processor versions.
If you have any questions or require assistance, contact us at Google Cloud support.
February 16, 2026
The layout parser web interface is in Preview.
It supports a document view display for processed PDF files and supports
visualizing the document's parsed JSON, block layout, and
image or table annotation data in a user-friendly interface. Bounding box support
only exists for version processor pretrained-layout-parser-v1.0-2024-06-03.
It also supports modifying the input layout config to allow for configuring the table and image annotation feature directly from the interface.
February 09, 2026
Layout parser model
pretrained-layout-parser-v1.6-2026-01-13 powered by Gemini 3
Flash LLM is available in Preview.
This processor version has ML processing capabilities in the US and EU.
For more information about available models, see the custom extractor page.
January 27, 2026
Layout parser model
pretrained-layout-parser-v1.6-pro-2025-12-01 powered by Gemini 3 Pro
LLM is available in Preview.
This processor version has ML processing capabilities in the US and EU.
For more information about available models, see the layout parser page.
Custom extractor model
pretrained-foundation-model-v1.6-pro-2025-12-01 powered by Gemini 3 Pro
LLM is available in Preview.
This processor version has ML processing capabilities in the US and EU.
For more information about available models, see the custom extractor page.
Custom extractor model
pretrained-foundation-model-v1.6-2026-01-13 powered by Gemini 3
Flash LLM is available in Preview.
This processor version has ML processing capabilities in the US and EU.
For more information about available models, see the custom extractor page.
January 12, 2026
Document AI is introducing document-level prompting for custom document processors. This feature allows you to provide an overall description of the document to inject deep business knowledge into the model, leading to improved extraction quality.
Document-level prompting offers improved accuracy by giving the model necessary context for extraction at the document level. This allows for easily supplied general information, such as geographical limitations (for example, all the address fields are located in the USA), to guide the model.
For more details, refer to the documentation on custom extractor mechanisms and document-level prompting.
December 15, 2025
A monitoring dashboard web interface is available in Preview to monitor at the project and processor level.
You can monitor a number of metrics, such as number of successfully processed
pages and sync processing latency, across fields like location, processor_type,
and processor_id over time.
For more information, see monitoring dashboard.
November 12, 2025
Automated schema extraction for custom extractor processors is in Preview.
This feature allows you to automatically extract a document schema from a test document you supply. Then, you can approve or decline the schema and edit it manually. This saves time and effort when defining the document schema for your custom processor and allows you to focus on refining the schema.
When creating a custom extractor processor, find the Generate from document option in the Get started tab of the Google Cloud console.
November 07, 2025
Gemini layout parser is in
Preview.
The Gemini layout parser gives better layout quality on table recognition,
reading order and text recognition on PDF files. You can enable the feature by
default by selecting layout parser processor version
pretrained-layout-parser-v1.4-2024-08-25, pretrained-layout-parser-v1.5-2025-08-25
or pretrained-layout-parser-v1.5-pro-2025-08-25 for your processor.
November 04, 2025
Layout parser support for DOCX, PPTX, XLSX, and XLSM file types in Document AI is in General Availability (GA). It makes content like paragraphs, tables, lists, and structural elements like headings, page headers, and footers easily accessible. It also creates context-aware chunks that facilitate information retrieval in a range of generative AI and discovery applications.
For more information, see Process documents with Layout Parser.
October 31, 2025
Custom splitter model
pretrained-splitter-v1.5-2025-07-14 with zero-shot splitting, classification
and confidence scores is available as Release Candidate
(Preview).
October 17, 2025
Layout parser lets you parse images and tables as annotations in Preview.
Layout parser can identify if there are images or tables in parsed documents. When found, images and tables are annotated as a descriptive block of text with the information depicted in the image and table.
October 06, 2025
Capacity reservation is available for Document AI in Preview. This lets you grant capacity to selected processors and maintain a steady real-time, high-volume processing flow for document processing requests.
For the necessary steps, read the Make a capacity reservation request section of "Quotas".
Custom extractor model
pretrained-foundation-model-v1.5.1-2025-08-07 with improved adaptive few-shot
learning is available as Release Candidate
(Preview).
Support for confidence scores in Custom classifier
models pretrained-foundation-model-v1.4-2025-05-16 and pretrained-classifier-v1.5-2025-08-05
is in Preview.
For best performance, use them with fine-tuned models.
September 23, 2025
Custom classifier
model pretrained-classifier-v1.5-2025-08-05
powered by Gemini 2.5 Flash is in Preview. It has ML processing available for US and EU regions, a
maximum page limit of 30 pages,
and processing requests of 120 pages per minute.
Unlike the prior custom classifier, which used classical machine learning, this version features a new platform. It accommodates:
- High accuracy immediately, based on the document classes you define.
- Few-shot learning to further improve accuracy.
- Use of descriptions when labeling for more context and insight for document classes.
- More accurate results with the same training dataset on the fine-tuned generative AI model, compared to the trained version.
- Autolabeling documents for fine-tuning and evaluation.
- Generative AI to fine-tune and heighten accuracy.
For more information on processor versions, see Managing processor versions.
September 10, 2025
Custom Extractor version pretrained-foundation-model-v1.4-2025-02-05 will no longer be accessible on February 5, 2026.
To avoid service disruptions, migrate to a later version such as
pretrained-foundation-model-v1.5-2025-05-05 or pretrained-foundation-model-v1.5-pro-2025-06-20.
To learn more about the migration process, refer to Manage processor
versions.
September 09, 2025
Document AI supports two service tiers and associated quotas: provisioned and best effort tiers.
The base is the provisioned tier quota, which provides 120 pages per minute for Gemini 2.0 and 2.5 Flash LLM and 30 pages per minute for Gemini 2.5 Pro LLM.
If you require more volume, best effort tier quota provides 120 pages per
minute for Gemini 2.0 2.5 Flash and 60 pages per minute for
Gemini 2.5 Pro. It's only used when the provisioned quota has been
exhausted. This applies to the BestEffortOnlineProcessDocumentPagesPerMinutePerProjectUS
and EU quotas and, in the console, best_effort_online_process_document_pages_us and eu.
Best effort can get up to 240 pages per minute for custom data extractor models v1.4 and v1.5 with a quota increase request (QIR). You can make a QIR by contacting your sales team representative.
There is no service level agreement (SLA) for best effort tier.
September 03, 2025
Custom extractor model pretrained-foundation-model-v1.5-pro-2025-06-20 is
available as General Availability (GA).
For more information about available models, see the custom extractor page.
August 29, 2025
Derived entity and signature detection are now supported in custom
extractor models pretrained-foundation-model-v1.4-2025-02-05
as General Availability (GA)
and in pretrained-foundation-model-v1.5-2025-05-05, as well as pretrained-foundation-model-v1.5-pro-2025-06-20
as Preview.
Signature detection lets you identify handwritten signatures by using visual cues in the document. Derived entity detection lets you deduce entities by inference without requiring the value to be explicitly present in the text. You can use this feature to deduce the country in an address, counting items in a table, or detecting if an ID is fake.
These can be enabled in the console when creating labels or by using the
DocumentSchema.EntityType
resource in the API.
For more information, read Custom extractor with derived fields and choose label attributes.
July 22, 2025
Custom extractor model
pretrained-foundation-model-v1.5-pro-2025-06-20
powered by Gemini 2.5 Pro is in Preview.
It has ML processing available for US and EU regions, a maximum page limit of
30 pages, and processing requests
of 30 pages per minute.
For more information, see Managing processor versions.
July 04, 2025
Document AI VPC service controls (VPC-SC) integration now supports identity groups.
For more information on setting up VPC-SC identity groups, read Configure identity groups and third-party identities in ingress and egress rules.
Document AI now supports Identity and Access Management (IAM) deny policies. These policies allow you to define deny rules that prevent certain principals from using certain permissions to access Google Cloud resources, regardless of the roles they're granted.
For more information, read Deny policy overview and Document AI security and compliance.
July 03, 2025
The Document AI CDE processor now supports merging the child entities
of nested entities that extend across several pages. This is supported in custom
extractor model pretrained-foundation-model-v1.5-2025-05-05.
This change is automatic in all processors.
For customers with existing v1.5 processors, to make use of this feature, you must relabel the nested entities in different pages.
To learn more about the labeling process, refer to Label documents.
June 30, 2025
Custom Extractor model pretrained-foundation-model-v1.5-2025-05-05 is in General Availability (GA) and has fine-tuning available for the US and EU.
From version v1.4 and later, we will use a new quota for online processing called Number of online process document pages per minute per processor type and model version. This quota will be enforced at a per-page and per-foundation model level. There will be no change to the batch processing quota.
These can be enabled in the console when creating labels and by using the DocumentSchema.EntityType.
For more information, read Managing processor versions.
June 19, 2025
We've increased the maximum file size for online processing requests from 20 MB to 40 MB. This applies to all types of processors.
For more information, see the Document AI limits page.
May 19, 2025
Cross-regions importing of fine-tuned models is now supported for processor versions based on Gemini 1.5 and later, such as
custom extractors
pretrained-foundation-model-v1.2-2024-05-10 and later.
For more information, see Managing processor versions.
May 05, 2025
Custom extractor model pretrained-foundation-model-v1.5-2025-04-25 powered by Gemini 2.5 Flash LLM is available as Public Preview in US regions. The custom extractor model supports a quota of up to 15 pages per minute for online process requests.
For more information about available models, see Custom extractor model versions.
April 08, 2025
Previous Custom Extractor versions pretrained-foundation-model-v1.0-2023-08-22 and pretrained-foundation-model-v1.1-2024-03-12 will be deprecated on April 9, 2025. To ensure uninterrupted service, prediction traffic to these versions, including any fine-tuned variants, will be automatically redirected to the latest version, pretrained-foundation-model-v1.4-2025-02-05.
For guidance on how to fine-tune a new version, refer to the fine tuning documentation.
April 02, 2025
All processors can now extend the Maximum page limit for online and synchronous requests up to 30 pages.
To do so, enable imageless_mode in ProcessRequest.
For Custom Extractor, you will need to first request to be allowlisted for this feature by filling out the form Allowlist Request for 30 Page limit in CDE.
March 24, 2025
As we launch Custom Extractor version pretrained-foundation-model-v1.4-2025-02-05 in GA with fine tuning (in Preview), these versions will no longer be accessible effective September 24, 2025:
pretrained-foundation-model-v1.2-2024-05-10pretrained-foundation-model-v1.3-2024-08-31
To avoid service disruptions, migrate to a later version, such as