Document AI release notes

This page documents production updates to Document AI. We recommend that Document AI developers periodically check this list for any new announcements.

You can see the latest product updates for all of Google Cloud on the Google Cloud page, browse and filter all release notes in the Google Cloud console, or programmatically access release notes in BigQuery.

To get the latest product updates delivered to you, add the URL of this page to your feed reader, or add the feed URL directly.

July 17, 2026

v1 & v1beta3
Feature

Custom extractor model pretrained-foundation-model-v3.5-2026-05-26 powered by Gemini 3.5 Flash LLM is available in Preview.

This processor version has ML processing capabilities in the US and EU.

For more information about available models, see the custom extractor page.

June 04, 2026

v1 & v1beta3
Feature

Custom extractor offers document validation and correction in Preview.

This feature allows you to enhance extraction accuracy with validation rules and document data using Common Expression Language (CEL) dialect.

For more information, see CEL dialect for document validation.

May 27, 2026

v1 & v1beta3
Feature

Layout parser image and table annotations is in General Availability (GA).

Layout parser can identify if there are images or tables in parsed documents. When found, images and tables are annotated as a descriptive block of text with the information depicted in the image and table.

March 31, 2026

v1beta3
Feature

Upgrading fine tuned custom extractor processors is now available in Preview.

The feature allows you to fine tune a new processor version with a newer base version, while keeping the configurations of the previously fine-tuned processor version selected. This is available through the UI in the Deploy & use tab in the console.

This is currently supported for upgrading pretrained-foundation-model-v1.4-2025-02-05 to pretrained-foundation-model-v1.5-2025-05-05.

For more information, see training overview.

March 27, 2026

v1 & v1beta3
Feature

Custom splitter model pretrained-splitter-v1.5-2025-07-14 is available in General Availability (GA).

March 23, 2026

v1 & v1beta3
Feature

Custom classifier models pretrained-classifier-v1.6-2026-03-09 and pretrained-classifier-v1.6-pro-2026-03-09 are available in Preview.

v1 & v1beta3
Feature

Custom splitter models pretrained-splitter-v1.6-2026-03-09 and pretrained-splitter-v1.6-pro-2026-03-09 are available in Preview.

March 03, 2026

v1 & v1beta3
Feature

Custom classifier model pretrained-classifier-v1.5-2025-08-05 is available as General Availability (GA).

For more information about available models, see the custom classifier page.

February 17, 2026

v1 & v1beta3
Deprecated

Document AI legacy processors will be discontinued on June 30, 2026. To preempt the risk of service failure while using legacy processors, we recommend transitioning to more stable, higher-quality processors.

The affected versions are:

Type Version
Identity parsers pretrained-us-passport-v1.0-2021-06-14
pretrained-fr-driver-license-v1.0-2021-06-14
Tax and finance parsers pretrained-1099misc-v1.1-2021-12-10
pretrained-1099nec-v1.0-2021-08-11
pretrained-1099r-v2.0-2022-07-25
pretrained-1099int-v1.1-2021-12-10
pretrained-ssa1099-v1.0-2021-08-09
pretrained-1099g-v1.0-2021-05-27
pretrained-1099g-v1.1-2021-12-10
pretrained-1120-v3.0-2022-04-26
pretrained-w9-v1.0-2020-09-25
pretrained-w9-v1.1-2021-12-10
pretrained-w9-v1.2-2022-01-27
pretrained-w9-v2.0-2022-06-23
Mortage and banking parsers pretrained-mortgage-statement-v1.0-2021-10-17
Procurement pretrained-utility-v1.1-2021-04-09
pretrained-utility-v1.2-2022-12-15
Splitting pretrained-procurement-splitter-v1.1-2021-04-09
pretrained-procurement-splitter-v1.2-2022-08-19
pretrained-lending-document-split-v1.0-2021-12-08
pretrained-lending-document-split-v2.0-2021-12-09
Summary pretrained-foundation-model-v1.0-2023-08-22

To ensure uninterrupted service and benefit from improved extraction quality, we recommend you migrate to the following later versions before June 30, 2026:

To learn more about the migration process, refer to Manage processor versions.

If you have any questions or require assistance, contact us at Google Cloud support.

February 16, 2026

v1 & v1beta3
Feature

The layout parser web interface is in Preview.

It supports a document view display for processed PDF files and supports visualizing the document's parsed JSON, block layout, and image or table annotation data in a user-friendly interface. Bounding box support only exists for version processor pretrained-layout-parser-v1.0-2024-06-03.

It also supports modifying the input layout config to allow for configuring the table and image annotation feature directly from the interface.

February 09, 2026

v1 & v1beta3
Feature

Layout parser model pretrained-layout-parser-v1.6-2026-01-13 powered by Gemini 3 Flash LLM is available in Preview.

This processor version has ML processing capabilities in the US and EU.

For more information about available models, see the custom extractor page.

January 27, 2026

v1 & v1beta3
Feature

Layout parser model pretrained-layout-parser-v1.6-pro-2025-12-01 powered by Gemini 3 Pro LLM is available in Preview.

This processor version has ML processing capabilities in the US and EU.

For more information about available models, see the layout parser page.

v1 & v1beta3
Feature

Custom extractor model pretrained-foundation-model-v1.6-pro-2025-12-01 powered by Gemini 3 Pro LLM is available in Preview.

This processor version has ML processing capabilities in the US and EU.

For more information about available models, see the custom extractor page.

v1 & v1beta3
Feature

Custom extractor model pretrained-foundation-model-v1.6-2026-01-13 powered by Gemini 3 Flash LLM is available in Preview.

This processor version has ML processing capabilities in the US and EU.

For more information about available models, see the custom extractor page.

January 12, 2026

v1 & v1beta3
Feature

Document AI is introducing document-level prompting for custom document processors. This feature allows you to provide an overall description of the document to inject deep business knowledge into the model, leading to improved extraction quality.

Document-level prompting offers improved accuracy by giving the model necessary context for extraction at the document level. This allows for easily supplied general information, such as geographical limitations (for example, all the address fields are located in the USA), to guide the model.

For more details, refer to the documentation on custom extractor mechanisms and document-level prompting.

December 15, 2025

v1 & v1beta3
Feature

A monitoring dashboard web interface is available in Preview to monitor at the project and processor level.

You can monitor a number of metrics, such as number of successfully processed pages and sync processing latency, across fields like location, processor_type, and processor_id over time.

For more information, see monitoring dashboard.

November 12, 2025

v1beta3
Feature

Automated schema extraction for custom extractor processors is in Preview.

This feature allows you to automatically extract a document schema from a test document you supply. Then, you can approve or decline the schema and edit it manually. This saves time and effort when defining the document schema for your custom processor and allows you to focus on refining the schema.

When creating a custom extractor processor, find the Generate from document option in the Get started tab of the Google Cloud console.

November 07, 2025

v1 & v1beta3
Feature

Gemini layout parser is in Preview. The Gemini layout parser gives better layout quality on table recognition, reading order and text recognition on PDF files. You can enable the feature by default by selecting layout parser processor version pretrained-layout-parser-v1.4-2024-08-25, pretrained-layout-parser-v1.5-2025-08-25 or pretrained-layout-parser-v1.5-pro-2025-08-25 for your processor.

November 04, 2025

v1 & v1beta3
Feature

Layout parser support for DOCX, PPTX, XLSX, and XLSM file types in Document AI is in General Availability (GA). It makes content like paragraphs, tables, lists, and structural elements like headings, page headers, and footers easily accessible. It also creates context-aware chunks that facilitate information retrieval in a range of generative AI and discovery applications.

For more information, see Process documents with Layout Parser.

October 31, 2025

v1 & v1beta3
Feature

Custom splitter model pretrained-splitter-v1.5-2025-07-14 with zero-shot splitting, classification and confidence scores is available as Release Candidate (Preview).

October 17, 2025

v1 & v1beta3
Feature

Layout parser lets you parse images and tables as annotations in Preview.

Layout parser can identify if there are images or tables in parsed documents. When found, images and tables are annotated as a descriptive block of text with the information depicted in the image and table.

October 06, 2025

v1 & v1beta3
Announcement

Capacity reservation is available for Document AI in Preview. This lets you grant capacity to selected processors and maintain a steady real-time, high-volume processing flow for document processing requests.

For the necessary steps, read the Make a capacity reservation request section of "Quotas".

v1 & v1beta3
Feature

Custom extractor model pretrained-foundation-model-v1.5.1-2025-08-07 with improved adaptive few-shot learning is available as Release Candidate (Preview).

v1 & v1beta3
Feature

Support for confidence scores in Custom classifier models pretrained-foundation-model-v1.4-2025-05-16 and pretrained-classifier-v1.5-2025-08-05 is in Preview.

For best performance, use them with fine-tuned models.

September 23, 2025

v1 & v1beta3
Feature

Custom classifier model pretrained-classifier-v1.5-2025-08-05 powered by Gemini 2.5 Flash is in Preview. It has ML processing available for US and EU regions, a maximum page limit of 30 pages, and processing requests of 120 pages per minute.

Unlike the prior custom classifier, which used classical machine learning, this version features a new platform. It accommodates:

  • High accuracy immediately, based on the document classes you define.
  • Few-shot learning to further improve accuracy.
  • Use of descriptions when labeling for more context and insight for document classes.
  • More accurate results with the same training dataset on the fine-tuned generative AI model, compared to the trained version.
  • Autolabeling documents for fine-tuning and evaluation.
  • Generative AI to fine-tune and heighten accuracy.

For more information on processor versions, see Managing processor versions.

September 10, 2025

v1 & v1beta3
Deprecated

Custom Extractor version pretrained-foundation-model-v1.4-2025-02-05 will no longer be accessible on February 5, 2026.

To avoid service disruptions, migrate to a later version such as pretrained-foundation-model-v1.5-2025-05-05 or pretrained-foundation-model-v1.5-pro-2025-06-20. To learn more about the migration process, refer to Manage processor versions.

September 09, 2025

v1beta3 & v1
Announcement

Document AI supports two service tiers and associated quotas: provisioned and best effort tiers.

The base is the provisioned tier quota, which provides 120 pages per minute for Gemini 2.0 and 2.5 Flash LLM and 30 pages per minute for Gemini 2.5 Pro LLM.

If you require more volume, best effort tier quota provides 120 pages per minute for Gemini 2.0 2.5 Flash and 60 pages per minute for Gemini 2.5 Pro. It's only used when the provisioned quota has been exhausted. This applies to the BestEffortOnlineProcessDocumentPagesPerMinutePerProjectUS and EU quotas and, in the console, best_effort_online_process_document_pages_us and eu.

Best effort can get up to 240 pages per minute for custom data extractor models v1.4 and v1.5 with a quota increase request (QIR). You can make a QIR by contacting your sales team representative.

There is no service level agreement (SLA) for best effort tier.

September 03, 2025

v1beta3 & v1
Feature

Custom extractor model pretrained-foundation-model-v1.5-pro-2025-06-20 is available as General Availability (GA).

For more information about available models, see the custom extractor page.

August 29, 2025

v1
Feature

Derived entity and signature detection are now supported in custom extractor models pretrained-foundation-model-v1.4-2025-02-05 as General Availability (GA) and in pretrained-foundation-model-v1.5-2025-05-05, as well as pretrained-foundation-model-v1.5-pro-2025-06-20 as Preview.

Signature detection lets you identify handwritten signatures by using visual cues in the document. Derived entity detection lets you deduce entities by inference without requiring the value to be explicitly present in the text. You can use this feature to deduce the country in an address, counting items in a table, or detecting if an ID is fake.

These can be enabled in the console when creating labels or by using the DocumentSchema.EntityType resource in the API.

For more information, read Custom extractor with derived fields and choose label attributes.

July 22, 2025

v1 & v1beta3
Feature

Custom extractor model pretrained-foundation-model-v1.5-pro-2025-06-20 powered by Gemini 2.5 Pro is in Preview. It has ML processing available for US and EU regions, a maximum page limit of 30 pages, and processing requests of 30 pages per minute.

For more information, see Managing processor versions.

July 04, 2025

v1 & v1beta3
Feature

Document AI VPC service controls (VPC-SC) integration now supports identity groups.

For more information on setting up VPC-SC identity groups, read Configure identity groups and third-party identities in ingress and egress rules.

v1 & v1beta3
Feature

Document AI now supports Identity and Access Management (IAM) deny policies. These policies allow you to define deny rules that prevent certain principals from using certain permissions to access Google Cloud resources, regardless of the roles they're granted.

For more information, read Deny policy overview and Document AI security and compliance.

July 03, 2025

v1 & v1beta3
Feature

The Document AI CDE processor now supports merging the child entities of nested entities that extend across several pages. This is supported in custom extractor model pretrained-foundation-model-v1.5-2025-05-05.

This change is automatic in all processors.

For customers with existing v1.5 processors, to make use of this feature, you must relabel the nested entities in different pages.

To learn more about the labeling process, refer to Label documents.

June 30, 2025

v1 & v1beta3
Feature

Custom Extractor model pretrained-foundation-model-v1.5-2025-05-05 is in General Availability (GA) and has fine-tuning available for the US and EU.

From version v1.4 and later, we will use a new quota for online processing called Number of online process document pages per minute per processor type and model version. This quota will be enforced at a per-page and per-foundation model level. There will be no change to the batch processing quota.

These can be enabled in the console when creating labels and by using the DocumentSchema.EntityType.

For more information, read Managing processor versions.

June 19, 2025

v1 & v1beta3
Feature

We've increased the maximum file size for online processing requests from 20 MB to 40 MB. This applies to all types of processors.

For more information, see the Document AI limits page.

May 19, 2025

v1 & v1beta3
Feature

Cross-regions importing of fine-tuned models is now supported for processor versions based on Gemini 1.5 and later, such as custom extractors pretrained-foundation-model-v1.2-2024-05-10 and later.

For more information, see Managing processor versions.

May 05, 2025

v1 & v1beta3
Feature

Custom extractor model pretrained-foundation-model-v1.5-2025-04-25 powered by Gemini 2.5 Flash LLM is available as Public Preview in US regions. The custom extractor model supports a quota of up to 15 pages per minute for online process requests.

For more information about available models, see Custom extractor model versions.

April 08, 2025

v1
Announcement

Previous Custom Extractor versions pretrained-foundation-model-v1.0-2023-08-22 and pretrained-foundation-model-v1.1-2024-03-12 will be deprecated on April 9, 2025. To ensure uninterrupted service, prediction traffic to these versions, including any fine-tuned variants, will be automatically redirected to the latest version, pretrained-foundation-model-v1.4-2025-02-05.

For guidance on how to fine-tune a new version, refer to the fine tuning documentation.

April 02, 2025

v1 & v1beta3
Feature

All processors can now extend the Maximum page limit for online and synchronous requests up to 30 pages.

To do so, enable imageless_mode in ProcessRequest.

For Custom Extractor, you will need to first request to be allowlisted for this feature by filling out the form Allowlist Request for 30 Page limit in CDE.

March 24, 2025

v1
Deprecated

As we launch Custom Extractor version pretrained-foundation-model-v1.4-2025-02-05 in GA with fine tuning (in Preview), these versions will no longer be accessible effective September 24, 2025:

  • pretrained-foundation-model-v1.2-2024-05-10
  • pretrained-foundation-model-v1.3-2024-08-31

To avoid service disruptions, migrate to a later version, such as