AWS Certified AI PractitionerAI/ML and Generative AI FundamentalsEasy
A data scientist is preparing a dataset for an image classification task where the goal is to identify different species of birds. The dataset contains thousands of images, each with a corresponding label indicating the bird species. Which type of machine learning task does this scenario represent?
- ARegression
- BAnomaly Detection
- CClustering
- DClassification
Show answer & explanationAnswer & explanation
Correct answer: D. Classification
This scenario describes a classic classification task. The model is trained on labeled images (input features) to predict a discrete category (the bird species, which is the label or target variable).
Why the other options are wrong
- A. Regression is used for predicting continuous numerical values, not discrete categories like bird species.
- B. Anomaly detection is about identifying rare, unusual patterns or outliers, not categorizing known items into predefined classes.
- C. Clustering is an unsupervised learning technique for grouping data without prior labels; here, labels are explicitly provided.
Classification (ML Task)
Classification is a supervised machine learning task where the model learns from labeled data to predict a discrete category or class for new, unseen data points.
- Predicts discrete, distinct categories (e.g., 'spam'/'not spam', 'cat'/'dog').
- Requires labeled training data.
- Can be binary (two classes) or multi-class (more than two classes).
Memory trick: Labeled data leads to predictable outcomes.