Microsoft Azure AI Fundamentals (AI-900)Describe features of computer vision workloads on AzureMedium

A logistics company wants to automate the process of reading shipping labels from packages as they move along a conveyor belt. The labels contain alphanumeric tracking codes and destination addresses. Which computer vision capability is essential for this task?

  1. AImage Classification
  2. BOptical Character Recognition (OCR)
  3. CSemantic Segmentation
  4. DObject Detection
Show answer & explanation

Correct answer: B. Optical Character Recognition (OCR)

Optical Character Recognition (OCR) is specifically designed to extract text from images, making it essential for reading alphanumeric tracking codes and addresses from shipping labels.

Why the other options are wrong

  • A. Image Classification would categorize the entire package, not read text from its label.
  • C. Semantic Segmentation classifies each pixel into a category, which is useful for precise object boundaries but not for text extraction.
  • D. Object Detection would identify the label itself, but not extract the text content from it.

Optical Character Recognition (OCR)

The electronic or mechanical conversion of images of typed, handwritten or printed text into machine-encoded text.

  • Enables searchability, editing, and storage of text.
  • Used for digitizing documents, processing invoices, and license plate recognition.
  • Can handle various fonts, sizes, and orientations.

Memory trick: To read text from images, you need OCR.

More Describe features of computer vision workloads on Azure questions