Microsoft Certified: Azure AI Engineer AssociateImplement image and video processing solutionsMedium
A municipal government is developing an application to help visually impaired citizens navigate public spaces. The application needs to read text from street signs, bus schedules, and building directories in real-time, often under varying lighting conditions and angles. Which Azure AI service feature is best suited for this robust optical character recognition (OCR) requirement?
- AAzure Form Recognizer's Prebuilt Receipt model
- BAzure Computer Vision's Read API
- CAzure Custom Vision's Classification
- DAzure Computer Vision's General OCR
Show answer & explanationAnswer & explanation
Correct answer: B. Azure Computer Vision's Read API
The Azure Computer Vision Read API is specifically designed for highly accurate and robust OCR, capable of handling diverse text types, orientations, and challenging environmental conditions like varying lighting and angles. This makes it ideal for real-time text extraction from signs and directories.
Why the other options are wrong
- A. Form Recognizer's Prebuilt Receipt model is for extracting structured data from receipts, not for general real-time OCR from diverse public signs.
- C. Custom Vision Classification is for categorizing images, not for extracting text from them.
- D. General OCR is an older, less accurate OCR engine and would struggle with varying lighting, angles, and real-world complexities.
Azure Computer Vision Read API
An advanced OCR service that extracts printed and handwritten text from images and documents with high accuracy, even from challenging real-world scenarios.
- Robust against image quality issues (blur, glare, low light)
- Handles various text orientations and sizes
- Supports multiple languages
- Provides text lines, words, and bounding box information
Memory trick: Read API Reads Real-World Signs Reliably.