Microsoft Azure AI Fundamentals (AI-900)Describe features of computer vision workloads on AzureMedium

A research institution is analyzing historical documents, some of which are handwritten and others typewritten. They need a computer vision tool that can read and extract text from both types of documents, even if the handwriting is messy or the typewritten text is faded. Which Azure AI Computer Vision tool is specifically designed for this advanced text extraction?

  1. AAzure Computer Vision API (Image Tagging)
  2. BAzure Computer Vision API (Read API)
  3. CAzure Form Recognizer
  4. DAzure Custom Vision
Show answer & explanation

Correct answer: B. Azure Computer Vision API (Read API)

The Azure Computer Vision API's Read API is specifically optimized for extracting text from images, including both printed and handwritten text, with high accuracy even from challenging conditions like faded or messy documents. This makes it ideal for historical documents.

Why the other options are wrong

  • A. Image Tagging assigns keywords, not extracts text.
  • C. Azure Form Recognizer is specialized for extracting structured data from forms and documents with predefined layouts, which might not be suitable for unstructured historical documents with varying layouts and handwriting.
  • D. Azure Custom Vision is for training custom image classification or object detection models, not general text extraction.

Azure Computer Vision Read API

A specialized capability within the Azure Computer Vision service designed for highly accurate OCR, capable of extracting text from various types of images, including handwritten and printed documents, even with complex layouts or challenging conditions.

  • Optimized for both printed and handwritten text.
  • Handles various languages and orientations.
  • More robust than basic OCR for complex scenarios like historical documents.

Memory trick: Read API deciphers even the oldest scrolls.

More Describe features of computer vision workloads on Azure questions