Microsoft Azure AI Fundamentals (AI-900)Describe features of computer vision workloads on AzureMedium
A company is developing a system to help visually impaired individuals navigate their surroundings by providing real-time audio descriptions of what's in front of them. This requires the AI service to interpret the entire visual scene and generate a natural language sentence describing its content. Which Azure AI service provides this specific capability?
- AAzure Computer Vision - Image Captioning
- BAzure Form Recognizer
- CAzure Computer Vision - Read API
- DAzure Face API
Show answer & explanationAnswer & explanation
Correct answer: A. Azure Computer Vision - Image Captioning
Azure Computer Vision's Image Captioning capability is specifically designed to generate a natural language description of an image's content, which directly provides the 'audio descriptions of what's in front of them' functionality required.
Why the other options are wrong
- B. Azure Form Recognizer extracts structured data from documents, irrelevant to describing general visual surroundings.
- C. The Read API extracts text from images, which is not the same as describing the entire visual scene.
- D. Azure Face API is for detecting and analyzing human faces, which does not provide a general scene description.
Azure Computer Vision - Image Captioning
An Azure AI Computer Vision feature that generates human-readable sentences describing the content of an image.
- Provides a natural language summary of visual content.
- Useful for accessibility, content management, and SEO.
- Part of the broader Azure Computer Vision service.
Memory trick: Azure Computer Vision captions the scene.