Microsoft Azure AI Fundamentals (AI-900)Describe features of computer vision workloads on AzureMedium
A security company is developing a system to monitor restricted areas. They need to analyze video feeds to detect and track individuals who enter these zones, identifying their movement patterns over time. Which two computer vision capabilities are essential for this scenario?
- AFace Recognition and Image Captioning
- BImage Classification and Optical Character Recognition
- CSemantic Segmentation and Depth Estimation
- DObject Detection and Object Tracking
Show answer & explanationAnswer & explanation
Correct answer: D. Object Detection and Object Tracking
Object Detection is needed to identify individuals (objects) in each frame, and Object Tracking is then used to follow these detected individuals across multiple frames to understand their movement patterns over time.
Why the other options are wrong
- A. Face Recognition identifies specific individuals, but the question asks to 'detect and track individuals' generally, not necessarily identify them by name. Image Captioning generates descriptions, which is not needed for tracking.
- B. Image Classification would only label the entire scene, and OCR extracts text, neither addresses detecting and tracking individuals.
- C. Semantic Segmentation provides pixel-level classification, and Depth Estimation determines distance, neither directly addresses the core need to detect and track moving individuals.
Object Detection & Tracking
Object Detection identifies objects within an image or video frame, while Object Tracking follows the movement of these identified objects across a sequence of frames.
- Object Detection: Locates objects with bounding boxes.
- Object Tracking: Maintains object identity and monitors path over time.
- Often used together in video analysis for surveillance, traffic monitoring, and sports analytics.
Memory trick: Detect, then Track the moving objects.