Microsoft Certified: Azure AI Engineer AssociateImplement image and video processing solutionsEasy

A retail company is looking to enhance its online product catalog by automatically generating descriptive metadata for newly uploaded product images. This includes identifying general objects present (e.g., 'shoes', 'table'), providing broad conceptual tags (e.g., 'fashion', 'indoor'), and detecting human-perceivable attributes. Which Azure Computer Vision capability should they use?

  1. ARead API
  2. BAnalyze Image (Tags and Categories)
  3. CObject Detection
  4. DSpatial Analysis
Show answer & explanation

Correct answer: B. Analyze Image (Tags and Categories)

Azure Computer Vision's Analyze Image operation, specifically its Tagging and Categorization features, can automatically identify general objects, provide conceptual tags, and classify images into a hierarchical taxonomy, which is ideal for generating broad descriptive metadata.

Why the other options are wrong

  • A. Read API is for extracting text, not general object identification or tagging.
  • C. Object Detection identifies specific objects with bounding boxes, but 'Analyze Image' provides broader tags and categories.
  • D. Spatial Analysis is for analyzing people's movement in video streams, not static image metadata.

Computer Vision Analyze Image (Tags & Categories)

An Azure Computer Vision operation that returns a list of descriptive tags and categorizes an image into a hierarchical taxonomy based on its content.

  • Identifies general objects and concepts (tags)
  • Classifies images into a predefined category hierarchy
  • Provides confidence scores for tags and categories
  • Useful for content management and search

Memory trick: Analyze Image 'tags' and 'sorts' your pictures for better organization.

More Implement image and video processing solutions questions