Microsoft Certified: Azure AI Engineer AssociateImplement image and video processing solutionsHard
A local government agency needs to process a large volume of scanned historical documents, which include various official forms, handwritten notes, and printed reports. They require a solution to extract all textual content from these documents accurately, regardless of text orientation or quality, and then identify the overall language of each document. Which Azure AI service feature combination is most appropriate?
- AAzure Form Recognizer's Layout API and Azure Computer Vision's General OCR.
- BAzure Computer Vision's Read API for OCR and Azure Text Analytics for Language Detection.
- CAzure Computer Vision's Analyze Image and Azure Text Analytics for Key Phrase Extraction.
- DAzure Custom Vision for text classification and Azure Video Indexer.
Show answer & explanationAnswer & explanation
Correct answer: B. Azure Computer Vision's Read API for OCR and Azure Text Analytics for Language Detection.
The Azure Computer Vision Read API is highly effective for robust OCR on diverse documents, handling handwritten, printed, and varying orientations. Once the text is extracted, Azure Text Analytics' Language Detection feature can accurately identify the language of the textual content, fulfilling both requirements.
Why the other options are wrong
- A. Layout API is for structured data extraction, not primary OCR, and General OCR is less robust than Read API. This combination is less optimal for raw text extraction from diverse documents.
- C. Analyze Image provides general tags/description, not high-accuracy OCR. Key Phrase Extraction assumes the language is already known, not for language detection itself.
- D. Custom Vision is for image classification/object detection, not OCR. Video Indexer is for video analysis.
Computer Vision Read API + Text Analytics Language Detection
A powerful combination for extracting highly accurate text from diverse image documents using Read API, followed by identifying the language of the extracted text using Text Analytics.
- Read API handles complex OCR (handwritten, printed, mixed)
- Read API is robust to image quality issues and orientation
- Text Analytics detects the predominant language of given text
- Useful for digitizing and categorizing multilingual document archives
Memory trick: Read Text, Then Detect Its Language Easily.