AWS Certified AI PractitionerAI/ML and Generative AI FundamentalsEasy

A data scientist is preparing a dataset for an image classification task where the goal is to identify different species of birds. The dataset contains thousands of images, each with a corresponding label indicating the bird species. Which type of machine learning task does this scenario represent?

  1. ARegression
  2. BAnomaly Detection
  3. CClustering
  4. DClassification
Show answer & explanation

Correct answer: D. Classification

This scenario describes a classic classification task. The model is trained on labeled images (input features) to predict a discrete category (the bird species, which is the label or target variable).

Why the other options are wrong

  • A. Regression is used for predicting continuous numerical values, not discrete categories like bird species.
  • B. Anomaly detection is about identifying rare, unusual patterns or outliers, not categorizing known items into predefined classes.
  • C. Clustering is an unsupervised learning technique for grouping data without prior labels; here, labels are explicitly provided.

Classification (ML Task)

Classification is a supervised machine learning task where the model learns from labeled data to predict a discrete category or class for new, unseen data points.

  • Predicts discrete, distinct categories (e.g., 'spam'/'not spam', 'cat'/'dog').
  • Requires labeled training data.
  • Can be binary (two classes) or multi-class (more than two classes).

Memory trick: Labeled data leads to predictable outcomes.

More AI/ML and Generative AI Fundamentals questions