在线小批量文件注释

在线（同步）请求 - 在线注释请求（images:annotate 或 files:annotate）会立即将内嵌注释返回给用户。在线注释请求限制了您可以在单个请求中添加注释的文件数量。使用 images:annotate 请求时，您只能指定少量要添加注释的图片 (<=16)；使用 files:annotate 请求时，您只能指定单个文件，并指定该文件中要添加注释的少量页面 (<=5)。
离线（异步）请求 - 离线注释请求（images:asyncBatchAnnotate 或 files:asyncBatchAnnotate）会启动长时间运行的操作 (LRO)，并且不会立即向调用方返回响应。当 LRO 完成后，注释会作为文件存储在您指定的 Cloud Storage 存储桶中。通过 images:asyncBatchAnnotate 请求，您可以为每个请求指定最多 2000 张图片；与在线请求相比，files:asyncBatchAnnotate 请求允许您指定更大批量的文件，并且一次可以为每个文件指定更多页面（<=2000）进行注释。

Vision API 可为 Cloud Storage 中存储的 PDF、TIFF 或 GIF 文件中的多个页面或帧提供在线（即时）注释。

您可以请求对为每个文件选择的 5 个帧（GIF；“image/gif”）或页面（PDF；“application/pdf”或 TIFF；“image/tiff”）进行在线特征检测和注释。

此页面中的示例注释针对的是 DOCUMENT_TEXT_DETECTION，而在线小批量注释适用于所有 Vision 特征。

注意：Vision API 还支持离线异步 PDF/TIFF 文件注释（目前仅适用于 DOCUMENT_TEXT_DETECTION 特征类型）。

离线异步请求会将响应 JSON 文件返回到您的 Cloud Storage 存储桶中，且支持多达 2000 个页面的文件。如需了解详情，请参阅检测文件 (PDF/TIFF) 中的文本。

PDF 文件的前五页 — gs://cloud-samples-data/vision/document_understanding/custom_0773375000.pdf

限制

最多可以为 5 个页面添加注释。用户可以指定 5 个特定页面来添加注释。

身份验证

files:annotate 请求不支持 API 密钥。

设置您的 Google Cloud 项目和身份验证

如果您尚未创建 Google Cloud 项目，请立即创建。展开本部分可查看相关说明。

Sign in to your Google Cloud account. If you're new to Google Cloud, create an account to evaluate how our products perform in real-world scenarios. New customers also get $300 in free credits to run, test, and deploy workloads.

In the Google Cloud console, on the project selector page, select or create a Google Cloud project.

Go to project selector

Make sure that billing is enabled for your Google Cloud project.

Enable the Vision API.

Enable the API

Install the Google Cloud CLI.

To initialize the gcloud CLI, run the following command:

gcloud init

In the Google Cloud console, on the project selector page, select or create a Google Cloud project.

Go to project selector

Make sure that billing is enabled for your Google Cloud project.

Enable the Vision API.

Enable the API

Install the Google Cloud CLI.

To initialize the gcloud CLI, run the following command:

gcloud init

目前支持的特征类型

特征类型
`CROP_HINTS`	确定图片的建议剪裁区域顶点。
`DOCUMENT_TEXT_DETECTION`	对文档 (PDF/TIFF) 等包含密集文本的图片和包含手写内容的图片执行 OCR。`TEXT_DETECTION` 可用于包含稀疏文本的图片。如果同时存在 `DOCUMENT_TEXT_DETECTION` 和 `TEXT_DETECTION`，则优先考虑。
`FACE_DETECTION`	检测图片中的人脸。
`IMAGE_PROPERTIES`	计算一组图片属性，例如图片的主色。
`LABEL_DETECTION`	根据图片内容添加标签。
`LANDMARK_DETECTION`	检测图片中的地标。
`LOGO_DETECTION`	检测图片中的公司徽标。
`OBJECT_LOCALIZATION`	检测并提取图片中的多个对象。
`SAFE_SEARCH_DETECTION`	运行安全搜索可检测可能不安全的内容或不良内容。
`TEXT_DETECTION`	对图片中的文本执行光学字符识别 (OCR)。文本检测针对大型图片中的稀疏文本区域进行了优化。如果图片为文档 (PDF/TIFF)、包含密集文本或包含手写内容，请改用 `DOCUMENT_TEXT_DETECTION`。
`WEB_DETECTION`	检测图片中的新闻、事件或名人等主题实体，并借助强大的 Google 图片搜索在网络上查找相似的图片。

示例代码

您可以使用本地存储的文件发送注释请求，也可以使用 Cloud Storage 上存储的文件。

使用本地存储的文件

使用以下代码示例获取本地存储的文件的任何特征注释。

REST

如需对一小批文件执行在线 PDF/TIFF/GIF 特征检测，请发出 POST 请求并提供相应的请求正文：

在使用任何请求数据之前，请先进行以下替换：

BASE64_ENCODED_FILE：二进制文件数据的 base64 表示（ASCII 字符串）。此字符串应类似于以下字符串：
- JVBERi0xLjUNCiW1tbW1...ydHhyZWYNCjk5NzM2OQ0KJSVFT0Y=
如需了解详情，请参阅 base64 编码主题。
PROJECT_ID：您的 Google Cloud 项目 ID。

特定于字段的注意事项：

inputConfig.mimeType - 下列类型之一：“application/pdf”“image/tiff”或“image/gif”。
pages - 指定要执行特征检测的文件的特定页面。

HTTP 方法和网址：

POST https://vision.googleapis.com/v1/files:annotate

请求 JSON 正文：

{
  "requests": [
    {
      "inputConfig": {
        "content": "BASE64_ENCODED_FILE",
        "mimeType": "application/pdf"
      },
      "features": [
        {
          "type": "DOCUMENT_TEXT_DETECTION"
        }
      ],
      "pages": [
        1,2,3,4,5
      ]
    }
  ]
}

如需发送请求，请选择以下方式之一：

curl

注意：以下命令假定您已使用您的用户账号通过运行 gcloud init 或 gcloud auth login 登录 gcloud CLI，或者使用了 Cloud Shell，这会使您自动登录 gcloud CLI。您可以运行 gcloud auth list 来检查当前活跃的账号。

将请求正文保存在名为 request.json 的文件中，然后执行以下命令：

curl -X POST \
     -H "Authorization: Bearer $(gcloud auth print-access-token)" \
     -H "x-goog-user-project: PROJECT_ID" \
     -H "Content-Type: application/json; charset=utf-8" \
     -d @request.json \
     "https://vision.googleapis.com/v1/files:annotate"

PowerShell

注意：以下命令假定您已使用您的用户账号通过运行 gcloud init 或 gcloud auth login 登录 gcloud CLI。您可以运行 gcloud auth list 来检查当前活跃的账号。

将请求正文保存在名为 request.json 的文件中，然后执行以下命令：

$cred = gcloud auth print-access-token
$headers = @{ "Authorization" = "Bearer $cred"; "x-goog-user-project" = "PROJECT_ID" }

Invoke-WebRequest `
    -Method POST `
    -Headers $headers `
    -ContentType: "application/json; charset=utf-8" `
    -InFile request.json `
    -Uri "https://vision.googleapis.com/v1/files:annotate" | Select-Object -Expand Content

响应：

成功的 annotate 请求会立即返回 JSON 响应。

对于此特征 (DOCUMENT_TEXT_DETECTION)，JSON 响应与图片的文档文本检测请求的响应类似。响应包含按段落、字词和各符号划分的文本块的边界框。还会检测全文。响应还包含一个 context 字段，显示指定的 PDF 或 TIFF 的位置以及结果在文件中的页码。

以下响应 JSON 仅针对单个页面（第 2 页），为清楚起见，这里采用简写形式。

响应

{
  "responses": [
    {
      "responses": [
        {
          "fullTextAnnotation": {
            "pages": [
              {
                "property": {
                  "detectedLanguages": [
                    {
                      "languageCode": "en",
                      "confidence": 0.99
                    },
                    {
                      "languageCode": "pl",
                      "confidence": 0.01
                    }
                  ]
                },
                "width": 1342,
                "height": 2234,
                "blocks": [
                  {
                    "boundingBox": {
                      "vertices": [
                      ...
                      ]
                    },
                    "paragraphs": [
                      {
                        "boundingBox": {
                          "vertices": [
                          ...
                          ]
                        },
                        "words": [
                          {
                            "property": {
                              "detectedLanguages": [
                                {
                                  "languageCode": "en"
                                }
                              ]
                            },
                            "boundingBox": {
                              "vertices": [
                              ...
                              ]
                            },
                            "symbols": [
                              {
                                "property": {
                                  "detectedLanguages": [
                                    {
                                      "languageCode": "en"
                                    }
                                  ],
                                  "detectedBreak": {
                                    "type": "SPACE"
                                  }
                                },
                                "boundingBox": {
                                  "vertices": [
                                ...
                                  ]
                                },
                                "text": "#",
                                "confidence": 0.07
                              }
                            ],
                            "confidence": 0.07
                          },
                          ...
                    ],
                    "blockType": "TEXT",
                    "confidence": 0.88
                  },
                  ...
            ...
            "text": "# THE STATE OF TEXAS\n0\nOIL, GAS AND MINERAL LEASE\n
            COUNTY OF MAVERICK\nTHIS AGREEMENT made this 14 day of_June\n1954,
            between Norvel J. Chittim and his wife, Lieschen G. Chittim;\nMary
            Anne Chittim Parker, joined herein pro forma by her husband,\nJoseph
            Bright Parker; Dorothea Chittim Oppenheimer, joined herein\nji pro
            forma by her husband, Fred J. Oppenheimer; Tuleta Chittim\nWright,
            joined herein pro forma by her husband, Gilbert G. Wright,\nJr.;
            Gilbert G. Wright, III; Delă Wright White, joined herein pro\nforma
            by her husband, John H. White; Anne Wright Basse, joined\nherein
            pro forma by her husband, E. A. Basse, Jr.; Norvel J.\nChittim,
            Independent Executor and Trustee for Estate of Marstella\nChittim,
            Deceased; Mary Louise Roswell, joined herein pro forma by\nher
            husband, Charles M. 'Roswell; and James M. Chittim and his wife\n
            Thelma Neal Chittim; as LESSORS, and W. L. Scheig of San Antonio,\n
            Texas, as LESSEE,\n10\nW ITNESS ETH:\nLessors, in consideration of
            $10.00, cash in hand paid, i\nof the royalties herein provided,
            and of the agreement of Lessee\nherein contained, hereby grant,
            lease and let exclusively unto\nLessee the tracts of land
            hereinafter described for the purpose of\ntesting for mineral
            indications, and in such tests use the Seismo-\ngraph, Torsion
            Balance, Core Drill, or any other tools, machinery,\nequipment
            or explosive necessary and proper; and also prospecting,\ndrilling
            and mining for and producing oil, gas and other minerals i\n
            (except metallic minerals), laying pipe lines, building tanks,\n
            power stations, telephone lines and other structures thereon to\n
            produce, save, take care of, treat, transport and own said pro-\n
            ducts and housing its employees (Lessee to conduct its geophysical\n
            work in such manner as not to damage the buildings, water tanks\n
            or wells of Lessors, or the livestock of Lessors or Lessors' ten-\n
            ants, ) said lands being situated in Maverick, Zavalla and Dimmit\n
            Counties, Texas, to-wit:\n3 -1.\n"
          },
          "context": {
            "pageNumber": 2
          }
        }
      ]
    }
  ]
}

Java

在试用此示例之前，请按照Vision API 快速入门：使用客户端库中的 Java 设置说明进行操作。如需了解详情，请参阅 Vision API Java 参考文档。

import com.google.cloud.vision.v1.AnnotateFileRequest;
import com.google.cloud.vision.v1.AnnotateImageResponse;
import com.google.cloud.vision.v1.BatchAnnotateFilesRequest;
import com.google.cloud.vision.v1.BatchAnnotateFilesResponse;
import com.google.cloud.vision.v1.Block;
import com.google.cloud.vision.v1.Feature;
import com.google.cloud.vision.v1.ImageAnnotatorClient;
import com.google.cloud.vision.v1.InputConfig;
import com.google.cloud.vision.v1.Page;
import com.google.cloud.vision.v1.Paragraph;
import com.google.cloud.vision.v1.Symbol;
import com.google.cloud.vision.v1.Word;
import com.google.protobuf.ByteString;
import java.io.IOException;
import java.nio.file.Files;
import java.nio.file.Path;
import java.nio.file.Paths;

public class BatchAnnotateFiles {

  public static void batchAnnotateFiles() throws IOException {
    String filePath = "path/to/your/file.pdf";
    batchAnnotateFiles(filePath);
  }

  public static void batchAnnotateFiles(String filePath) throws IOException {
    // Initialize client that will be used to send requests. This client only needs to be created
    // once, and can be reused for multiple requests. After completing all of your requests, call
    // the "close" method on the client to safely clean up any remaining background resources.
    try (ImageAnnotatorClient imageAnnotatorClient = ImageAnnotatorClient.create()) {
      // You can send multiple files to be annotated, this sample demonstrates how to do this with
      // one file. If you want to use multiple files, you have to create a `AnnotateImageRequest`
      // object for each file that you want annotated.
      // First read the files contents
      Path path = Paths.get(filePath);
      byte[] data = Files.readAllBytes(path);
      ByteString content = ByteString.copyFrom(data);

      // Specify the input config with the file's contents and its type.
      // Supported mime_type: application/pdf, image/tiff, image/gif
      // https://cloud.google.com/vision/docs/reference/rpc/google.cloud.vision.v1#inputconfig
      InputConfig inputConfig =
          InputConfig.newBuilder().setMimeType("application/pdf").setContent(content).build();

      // Set the type of annotation you want to perform on the file
      // https://cloud.google.com/vision/docs/reference/rpc/google.cloud.vision.v1#google.cloud.vision.v1.Feature.Type
      Feature feature = Feature.newBuilder().setType(Feature.Type.DOCUMENT_TEXT_DETECTION).build();

      // Build the request object for that one file. Note: for additional file you have to create
      // additional `AnnotateFileRequest` objects and store them in a list to be used below.
      // Since we are sending a file of type `application/pdf`, we can use the `pages` field to
      // specify which pages to process. The service can process up to 5 pages per document file.
      // https://cloud.google.com/vision/docs/reference/rpc/google.cloud.vision.v1#google.cloud.vision.v1.AnnotateFileRequest
      AnnotateFileRequest fileRequest =
          AnnotateFileRequest.newBuilder()
              .setInputConfig(inputConfig)
              .addFeatures(feature)
              .addPages(1) // Process the first page
              .addPages(2) // Process the second page
              .addPages(-1) // Process the last page
              .build();

      // Add each `AnnotateFileRequest` object to the batch request.
      BatchAnnotateFilesRequest request =
          BatchAnnotateFilesRequest.newBuilder().addRequests(fileRequest).build();

      // Make the synchronous batch request.
      BatchAnnotateFilesResponse response = imageAnnotatorClient.batchAnnotateFiles(request);

      // Process the results, just get the first result, since only one file was sent in this
      // sample.
      for (AnnotateImageResponse imageResponse :
          response.getResponsesList().get(0).getResponsesList()) {
        System.out.format("Full text: %s%n", imageResponse.getFullTextAnnotation().getText());
        for (Page page : imageResponse.getFullTextAnnotation().getPagesList()) {
          for (Block block : page.getBlocksList()) {
            System.out.format("%nBlock confidence: %s%n", block.getConfidence());
            for (Paragraph par : block.getParagraphsList()) {
              System.out.format("\tParagraph confidence: %s%n", par.getConfidence());
              for (Word word : par.getWordsList()) {
                System.out.format("\t\tWord confidence: %s%n", word.getConfidence());
                for (Symbol symbol : word.getSymbolsList()) {
                  System.out.format(
                      "\t\t\tSymbol: %s, (confidence: %s)%n",
                      symbol.getText(), symbol.getConfidence());
                }
              }
            }
          }
        }
      }
    }
  }
}

Node.js

试用此示例之前，请按照《Vision 快速入门：使用客户端库》中的 Node.js 设置说明进行操作。如需了解详情，请参阅 Vision Node.js API 参考文档。

如需向 Vision 进行身份验证，请设置应用默认凭据。如需了解详情，请参阅为本地开发环境设置身份验证。

/**
 * TODO(developer): Uncomment these variables before running the sample.
 */
// const fileName = 'path/to/your/file.pdf';

// Imports the Google Cloud client libraries
const {ImageAnnotatorClient} = require('@google-cloud/vision').v1;
const fs = require('fs').promises;

// Instantiates a client
const client = new ImageAnnotatorClient();

// You can send multiple files to be annotated, this sample demonstrates how to do this with
// one file. If you want to use multiple files, you have to create a request object for each file that you want annotated.
async function batchAnnotateFiles() {
  // First Specify the input config with the file's path and its type.
  // Supported mime_type: application/pdf, image/tiff, image/gif
  // https://cloud.google.com/vision/docs/reference/rpc/google.cloud.vision.v1#inputconfig
  const inputConfig = {
    mimeType: 'application/pdf',
    content: await fs.readFile(fileName),
  };

  // Set the type of annotation you want to perform on the file
  // https://cloud.google.com/vision/docs/reference/rpc/google.cloud.vision.v1#google.cloud.vision.v1.Feature.Type
  const features = [{type: 'DOCUMENT_TEXT_DETECTION'}];

  // Build the request object for that one file. Note: for additional files you have to create
  // additional file request objects and store them in a list to be used below.
  // Since we are sending a file of type `application/pdf`, we can use the `pages` field to
  // specify which pages to process. The service can process up to 5 pages per document file.
  // https://cloud.google.com/vision/docs/reference/rpc/google.cloud.vision.v1#google.cloud.vision.v1.AnnotateFileRequest
  const fileRequest = {
    inputConfig: inputConfig,
    features: features,
    // Annotate the first two pages and the last one (max 5 pages)
    // First page starts at 1, and not 0. Last page is -1.
    pages: [1, 2, -1],
  };

  // Add each `AnnotateFileRequest` object to the batch request.
  const request = {
    requests: [fileRequest],
  };

  // Make the synchronous batch request.
  const [result] = await client.batchAnnotateFiles(request);

  // Process the results, just get the first result, since only one file was sent in this
  // sample.
  const responses = result.responses[0].responses;

  for (const response of responses) {
    console.log(`Full text: ${response.fullTextAnnotation.text}`);
    for (const page of response.fullTextAnnotation.pages) {
      for (const block of page.blocks) {
        console.log(`Block confidence: ${block.confidence}`);
        for (const paragraph of block.paragraphs) {
          console.log(` Paragraph confidence: ${paragraph.confidence}`);
          for (const word of paragraph.words) {
            const symbol_texts = word.symbols.map(symbol => symbol.text);
            const word_text = symbol_texts.join('');
            console.log(
              `  Word text: ${word_text} (confidence: ${word.confidence})`
            );
            for (const symbol of word.symbols) {
              console.log(
                `   Symbol: ${symbol.text} (confidence: ${symbol.confidence})`
              );
            }
          }
        }
      }
    }
  }
}

batchAnnotateFiles();

Python

试用此示例之前，请按照《Vision 快速入门：使用客户端库》中的 Python 设置说明进行操作。如需了解详情，请参阅 Vision Python API 参考文档。

如需向 Vision 进行身份验证，请设置应用默认凭据。如需了解详情，请参阅为本地开发环境设置身份验证。



from google.cloud import vision_v1


def sample_batch_annotate_files(file_path="path/to/your/document.pdf"):
    """Perform batch file annotation."""
    client = vision_v1.ImageAnnotatorClient()

    # Supported mime_type: application/pdf, image/tiff, image/gif
    mime_type = "application/pdf"
    with open(file_path, "rb") as f:
        content = f.read()
    input_config = {"mime_type": mime_type, "content": content}
    features = [{"type_": vision_v1.Feature.Type.DOCUMENT_TEXT_DETECTION}]

    # The service can process up to 5 pages per document file. Here we specify
    # the first, second, and last page of the document to be processed.
    pages = [1, 2, -1]
    requests = [{"input_config": input_config, "features": features, "pages": pages}]

    response = client.batch_annotate_files(requests=requests)
    for image_response in response.responses[0].responses:
        print(f"Full text: {image_response.full_text_annotation.text}")
        for page in image_response.full_text_annotation.pages:
            for block in page.blocks:
                print(f"\nBlock confidence: {block.confidence}")
                for par in block.paragraphs:
                    print(f"\tParagraph confidence: {par.confidence}")
                    for word in par.words:
                        print(f"\t\tWord confidence: {word.confidence}")
                        for symbol in word.symbols:
                            print(
                                "\t\t\tSymbol: {}, (confidence: {})".format(
                                    symbol.text, symbol.confidence
                                )
                            )

使用 Cloud Storage 上的文件

使用以下代码示例获取 Cloud Storage 文件的任何特征注释。

REST

如需对一小批文件执行在线 PDF/TIFF/GIF 特征检测，请发出 POST 请求并提供相应的请求正文：

在使用任何请求数据之前，请先进行以下替换：

CLOUD_STORAGE_FILE_URI：Cloud Storage 存储桶中有效文件 (PDF/TIFF) 的路径。您必须至少拥有该文件的读取权限。示例：
- ```
gs://cloud-samples-data/vision/document_understanding/custom_0773375000.pdf
```
PROJECT_ID：您的 Google Cloud 项目 ID。

特定于字段的注意事项：

inputConfig.mimeType - 下列类型之一：“application/pdf”“image/tiff”或“image/gif”。
pages - 指定要执行特征检测的文件的特定页面。

HTTP 方法和网址：

POST https://vision.googleapis.com/v1/files:annotate

请求 JSON 正文：

{
  "requests": [
    {
      "inputConfig": {
        "gcsSource": {
          "uri": "CLOUD_STORAGE_FILE_URI"
        },
        "mimeType": "application/pdf"
      },
      "features": [
        {
          "type": "DOCUMENT_TEXT_DETECTION"
        }
      ],
      "pages": [
        1,2,3,4,5
      ]
    }
  ]
}

如需发送请求，请选择以下方式之一：

curl

将请求正文保存在名为 request.json 的文件中，然后执行以下命令：

curl -X POST \
     -H "Authorization: Bearer $(gcloud auth print-access-token)" \
     -H "x-goog-user-project: PROJECT_ID" \
     -H "Content-Type: application/json; charset=utf-8" \
     -d @request.json \
     "https://vision.googleapis.com/v1/files:annotate"

PowerShell

注意：以下命令假定您已使用您的用户账号通过运行 gcloud init 或 gcloud auth login 登录 gcloud CLI。您可以运行 gcloud auth list 来检查当前活跃的账号。

将请求正文保存在名为 request.json 的文件中，然后执行以下命令：

$cred = gcloud auth print-access-token
$headers = @{ "Authorization" = "Bearer $cred"; "x-goog-user-project" = "PROJECT_ID" }

Invoke-WebRequest `
    -Method POST `
    -Headers $headers `
    -ContentType: "application/json; charset=utf-8" `
    -InFile request.json `
    -Uri "https://vision.googleapis.com/v1/files:annotate" | Select-Object -Expand Content

响应：

成功的 annotate 请求会立即返回 JSON 响应。

以下响应 JSON 仅针对单个页面（第 2 页），为清楚起见，这里采用简写形式。

响应

{
  "responses": [
    {
      "responses": [
        {
          "fullTextAnnotation": {
            "pages": [
              {
                "property": {
                  "detectedLanguages": [
                    {
                      "languageCode": "en",
                      "confidence": 0.99
                    },
                    {
                      "languageCode": "pl",
                      "confidence": 0.01
                    }
                  ]
                },
                "width": 1342,
                "height": 2234,
                "blocks": [
                  {
                    "boundingBox": {
                      "vertices": [
                      ...
                      ]
                    },
                    "paragraphs": [
                      {
                        "boundingBox": {
                          "vertices": [
                          ...
                          ]
                        },
                        "words": [
                          {
                            "property": {
                              "detectedLanguages": [
                                {
                                  "languageCode": "en"
                                }
                              ]
                            },
                            "boundingBox": {
                              "vertices": [
                              ...
                              ]
                            },
                            "symbols": [
                              {
                                "property": {
                                  "detectedLanguages": [
                                    {
                                      "languageCode": "en"
                                    }
                                  ],
                                  "detectedBreak": {
                                    "type": "SPACE"
                                  }
                                },
                                "boundingBox": {
                                  "vertices": [
                                ...
                                  ]
                                },
                                "text": "#",
                                "confidence": 0.07
                              }
                            ],
                            "confidence": 0.07
                          },
                          ...
                    ],
                    "blockType": "TEXT",
                    "confidence": 0.88
                  },
                  ...
            ...
            "text": "# THE STATE OF TEXAS\n0\nOIL, GAS AND MINERAL LEASE\n
            COUNTY OF MAVERICK\nTHIS AGREEMENT made this 14 day of_June\n1954,
            between Norvel J. Chittim and his wife, Lieschen G. Chittim;\nMary
            Anne Chittim Parker, joined herein pro forma by her husband,\nJoseph
            Bright Parker; Dorothea Chittim Oppenheimer, joined herein\nji pro
            forma by her husband, Fred J. Oppenheimer; Tuleta Chittim\nWright,
            joined herein pro forma by her husband, Gilbert G. Wright,\nJr.;
            Gilbert G. Wright, III; Delă Wright White, joined herein pro\nforma
            by her husband, John H. White; Anne Wright Basse, joined\nherein
            pro forma by her husband, E. A. Basse, Jr.; Norvel J.\nChittim,
            Independent Executor and Trustee for Estate of Marstella\nChittim,
            Deceased; Mary Louise Roswell, joined herein pro forma by\nher
            husband, Charles M. 'Roswell; and James M. Chittim and his wife\n
            Thelma Neal Chittim; as LESSORS, and W. L. Scheig of San Antonio,\n
            Texas, as LESSEE,\n10\nW ITNESS ETH:\nLessors, in consideration of
            $10.00, cash in hand paid, i\nof the royalties herein provided,
            and of the agreement of Lessee\nherein contained, hereby grant,
            lease and let exclusively unto\nLessee the tracts of land
            hereinafter described for the purpose of\ntesting for mineral
            indications, and in such tests use the Seismo-\ngraph, Torsion
            Balance, Core Drill, or any other tools, machinery,\nequipment
            or explosive necessary and proper; and also prospecting,\ndrilling
            and mining for and producing oil, gas and other minerals i\n
            (except metallic minerals), laying pipe lines, building tanks,\n
            power stations, telephone lines and other structures thereon to\n
            produce, save, take care of, treat, transport and own said pro-\n
            ducts and housing its employees (Lessee to conduct its geophysical\n
            work in such manner as not to damage the buildings, water tanks\n
            or wells of Lessors, or the livestock of Lessors or Lessors' ten-\n
            ants, ) said lands being situated in Maverick, Zavalla and Dimmit\n
            Counties, Texas, to-wit:\n3 -1.\n"
          },
          "context": {
            "pageNumber": 2
          }
        }
      ]
    }
  ]
}

Java

在试用此示例之前，请按照Vision API 快速入门：使用客户端库中的 Java 设置说明进行操作。如需了解详情，请参阅 Vision API Java 参考文档。

import com.google.cloud.vision.v1.AnnotateFileRequest;
import com.google.cloud.vision.v1.AnnotateImageResponse;
import com.google.cloud.vision.v1.BatchAnnotateFilesRequest;
import com.google.cloud.vision.v1.BatchAnnotateFilesResponse;
import com.google.cloud.vision.v1.Block;
import com.google.cloud.vision.v1.Feature;
import com.google.cloud.vision.v1.GcsSource;
import com.google.cloud.vision.v1.ImageAnnotatorClient;
import com.google.cloud.vision.v1.InputConfig;
import com.google.cloud.vision.v1.Page;
import com.google.cloud.vision.v1.Paragraph;
import com.google.cloud.vision.v1.Symbol;
import com.google.cloud.vision.v1.Word;
import java.io.IOException;

public class BatchAnnotateFilesGcs {

  public static void batchAnnotateFilesGcs() throws IOException {
    String gcsUri = "gs://cloud-samples-data/vision/document_understanding/kafka.pdf";
    batchAnnotateFilesGcs(gcsUri);
  }

  public static void batchAnnotateFilesGcs(String gcsUri) throws IOException {
    // Initialize client that will be used to send requests. This client only needs to be created
    // once, and can be reused for multiple requests. After completing all of your requests, call
    // the "close" method on the client to safely clean up any remaining background resources.
    try (ImageAnnotatorClient imageAnnotatorClient = ImageAnnotatorClient.create()) {
      // You can send multiple files to be annotated, this sample demonstrates how to do this with
      // one file. If you want to use multiple files, you have to create a `AnnotateImageRequest`
      // object for each file that you want annotated.
      // First specify where the vision api can find the image
      GcsSource gcsSource = GcsSource.newBuilder().setUri(gcsUri).build();

      // Specify the input config with the file's uri and its type.
      // Supported mime_type: application/pdf, image/tiff, image/gif
      // https://cloud.google.com/vision/docs/reference/rpc/google.cloud.vision.v1#inputconfig
      InputConfig inputConfig =
          InputConfig.newBuilder().setMimeType("application/pdf").setGcsSource(gcsSource).build();

      // Set the type of annotation you want to perform on the file
      // https://cloud.google.com/vision/docs/reference/rpc/google.cloud.vision.v1#google.cloud.vision.v1.Feature.Type
      Feature feature = Feature.newBuilder().setType(Feature.Type.DOCUMENT_TEXT_DETECTION).build();

      // Build the request object for that one file. Note: for additional file you have to create
      // additional `AnnotateFileRequest` objects and store them in a list to be used below.
      // Since we are sending a file of type `application/pdf`, we can use the `pages` field to
      // specify which pages to process. The service can process up to 5 pages per document file.
      // https://cloud.google.com/vision/docs/reference/rpc/google.cloud.vision.v1#google.cloud.vision.v1.AnnotateFileRequest
      AnnotateFileRequest fileRequest =
          AnnotateFileRequest.newBuilder()
              .setInputConfig(inputConfig)
              .addFeatures(feature)
              .addPages(1) // Process the first page
              .addPages(2) // Process the second page
              .addPages(-1) // Process the last page
              .build();

      // Add each `AnnotateFileRequest` object to the batch request.
      BatchAnnotateFilesRequest request =
          BatchAnnotateFilesRequest.newBuilder().addRequests(fileRequest).build();

      // Make the synchronous batch request.
      BatchAnnotateFilesResponse response = imageAnnotatorClient.batchAnnotateFiles(request);

      // Process the results, just get the first result, since only one file was sent in this
      // sample.
      for (AnnotateImageResponse imageResponse :
          response.getResponsesList().get(0).getResponsesList()) {
        System.out.format("Full text: %s%n", imageResponse.getFullTextAnnotation().getText());
        for (Page page : imageResponse.getFullTextAnnotation().getPagesList()) {
          for (Block block : page.getBlocksList()) {
            System.out.format("%nBlock confidence: %s%n", block.getConfidence());
            for (Paragraph par : block.getParagraphsList()) {
              System.out.format("\tParagraph confidence: %s%n", par.getConfidence());
              for (Word word : par.getWordsList()) {
                System.out.format("\t\tWord confidence: %s%n", word.getConfidence());
                for (Symbol symbol : word.getSymbolsList()) {
                  System.out.format(
                      "\t\t\tSymbol: %s, (confidence: %s)%n",
                      symbol.getText(), symbol.getConfidence());
                }
              }
            }
          }
        }
      }
    }
  }
}

Node.js

试用此示例之前，请按照《Vision 快速入门：使用客户端库》中的 Node.js 设置说明进行操作。如需了解详情，请参阅 Vision Node.js API 参考文档。

如需向 Vision 进行身份验证，请设置应用默认凭据。如需了解详情，请参阅为本地开发环境设置身份验证。

/**
 * TODO(developer): Uncomment these variables before running the sample.
 */
// const gcsSourceUri = 'gs://cloud-samples-data/vision/document_understanding/kafka.pdf';

// Imports the Google Cloud client libraries
const {ImageAnnotatorClient} = require('@google-cloud/vision').v1;

// Instantiates a client
const client = new ImageAnnotatorClient();

// You can send multiple files to be annotated, this sample demonstrates how to do this with
// one file. If you want to use multiple files, you have to create a request object for each file that you want annotated.
async function batchAnnotateFiles() {
  // First Specify the input config with the file's uri and its type.
  // Supported mime_type: application/pdf, image/tiff, image/gif
  // https://cloud.google.com/vision/docs/reference/rpc/google.cloud.vision.v1#inputconfig
  const inputConfig = {
    mimeType: 'application/pdf',
    gcsSource: {
      uri: gcsSourceUri,
    },
  };

  // Set the type of annotation you want to perform on the file
  // https://cloud.google.com/vision/docs/reference/rpc/google.cloud.vision.v1#google.cloud.vision.v1.Feature.Type
  const features = [{type: 'DOCUMENT_TEXT_DETECTION'}];

  // Build the request object for that one file. Note: for additional files you have to create
  // additional file request objects and store them in a list to be used below.
  // Since we are sending a file of type `application/pdf`, we can use the `pages` field to
  // specify which pages to process. The service can process up to 5 pages per document file.
  // https://cloud.google.com/vision/docs/reference/rpc/google.cloud.vision.v1#google.cloud.vision.v1.AnnotateFileRequest
  const fileRequest = {
    inputConfig: inputConfig,
    features: features,
    // Annotate the first two pages and the last one (max 5 pages)
    // First page starts at 1, and not 0. Last page is -1.
    pages: [1, 2, -1],
  };

  // Add each `AnnotateFileRequest` object to the batch request.
  const request = {
    requests: [fileRequest],
  };

  // Make the synchronous batch request.
  const [result] = await client.batchAnnotateFiles(request);

  // Process the results, just get the first result, since only one file was sent in this
  // sample.
  const responses = result.responses[0].responses;

  for (const response of responses) {
    console.log(`Full text: ${response.fullTextAnnotation.text}`);
    for (const page of response.fullTextAnnotation.pages) {
      for (const block of page.blocks) {
        console.log(`Block confidence: ${block.confidence}`);
        for (const paragraph of block.paragraphs) {
          console.log(` Paragraph confidence: ${paragraph.confidence}`);
          for (const word of paragraph.words) {
            const symbol_texts = word.symbols.map(symbol => symbol.text);
            const word_text = symbol_texts.join('');
            console.log(
              `  Word text: ${word_text} (confidence: ${word.confidence})`
            );
            for (const symbol of word.symbols) {
              console.log(
                `   Symbol: ${symbol.text} (confidence: ${symbol.confidence})`
              );
            }
          }
        }
      }
    }
  }
}

batchAnnotateFiles();

Python

试用此示例之前，请按照《Vision 快速入门：使用客户端库》中的 Python 设置说明进行操作。如需了解详情，请参阅 Vision Python API 参考文档。

如需向 Vision 进行身份验证，请设置应用默认凭据。如需了解详情，请参阅为本地开发环境设置身份验证。


from google.cloud import vision_v1


def sample_batch_annotate_files(
    storage_uri="gs://cloud-samples-data/vision/document_understanding/kafka.pdf",
):
    """Perform batch file annotation."""
    mime_type = "application/pdf"

    client = vision_v1.ImageAnnotatorClient()

    gcs_source = {"uri": storage_uri}
    input_config = {"gcs_source": gcs_source, "mime_type": mime_type}
    features = [{"type_": vision_v1.Feature.Type.DOCUMENT_TEXT_DETECTION}]

    # The service can process up to 5 pages per document file.
    # Here we specify the first, second, and last page of the document to be
    # processed.
    pages = [1, 2, -1]
    requests = [{"input_config": input_config, "features": features, "pages": pages}]

    response = client.batch_annotate_files(requests=requests)
    for image_response in response.responses[0].responses:
        print(f"Full text: {image_response.full_text_annotation.text}")
        for page in image_response.full_text_annotation.pages:
            for block in page.blocks:
                print(f"\nBlock confidence: {block.confidence}")
                for par in block.paragraphs:
                    print(f"\tParagraph confidence: {par.confidence}")
                    for word in par.words:
                        print(f"\t\tWord confidence: {word.confidence}")
                        for symbol in word.symbols:
                            print(
                                "\t\t\tSymbol: {}, (confidence: {})".format(
                                    symbol.text, symbol.confidence
                                )
                            )

试用

请尝试下面的小批量在线特征检测。

您可以使用已指定的 PDF 文件，也可以指定自己的文件。

对于此请求，指定了三种特征类型：

DOCUMENT_TEXT_DETECTION
LABEL_DETECTION
CROP_HINTS

您可以通过更改请求中的相应对象 ({"type": "FEATURE_NAME"}) 来添加或移除其他特征类型。

选择执行即可发送请求。

请求正文：

{
  "requests": [
    {
      "inputConfig": {
        "gcsSource": {
          "uri": "gs://cloud-samples-data/vision/document_understanding/custom_0773375000.pdf"
        },
        "mimeType": "application/pdf"
      },
      "features": [
        {
          "type": "DOCUMENT_TEXT_DETECTION"
        },
        {
          "type": "LABEL_DETECTION"
        },
        {
          "type": "CROP_HINTS"
        }
      ],
      "pages": [
        1,
        2,
        3,
        4,
        5
      ]
    }
  ]
}

在线小批量文件注释 使用集合让一切井井有条 根据您的偏好保存内容并对其进行分类。

限制

身份验证

设置您的 Google Cloud 项目和身份验证

目前支持的特征类型

示例代码

使用本地存储的文件

REST

curl

PowerShell

响应

Java

Node.js

Python

使用 Cloud Storage 上的文件

REST

curl

PowerShell

响应

Java

Node.js

Python

试用

在线小批量文件注释