AWS Certified AI Practitioner practice questions
218 free questions with answers and explanations.
- 101.A research institution is developing a foundation model for scientific discovery, specifically focusing on bioinformatics. The model needs to process and interpret complex data types including DNA sequences (text), protein structures (3D graphs/images), and scientific literature (natural language text) simultaneously to identify new drug targets. Which type of foundation model is best suited for this requirement?Foundation Models
- 102.A manufacturing company wants to implement predictive maintenance for its industrial machinery. They have historical sensor data (temperature, vibration, pressure) and maintenance logs, but lack in-house machine learning expertise to build and train complex models. They need a fully managed service that can automatically build, train, and deploy ML models to predict equipment failures using their sensor data, without requiring them to write any code or manage ML infrastructure. Which AWS service is best suited for this task?AWS Services for AI/ML and Generative AI
- 103.A small e-commerce business wants to automatically generate unique and engaging product descriptions for thousands of items based on a few key features. They need a cost-effective solution that can handle a large volume of text generation without requiring extensive machine learning expertise. Which AWS service is best suited for this task?AWS Services for AI/ML and Generative AI
- 104.A development team is building a new application and wants to integrate an AI-powered coding assistant to help improve developer productivity by generating code suggestions and identifying vulnerabilities in real time. Which AWS service is designed to fulfill this requirement?AWS Services for AI/ML and Generative AI
- 105.A product manager is evaluating the use of foundation models for a new feature that automates customer email responses. The goal is to ensure the generated responses are helpful, harmless, and align with the company's brand voice and ethical guidelines. Which process is primarily focused on achieving these objectives by guiding the model's behavior towards desired human values and safety standards?Foundation Models
- 106.A healthcare provider wants to use a foundation model to assist doctors in diagnosing rare diseases by analyzing medical images, patient records, and genomic data. The primary concern is ensuring the model's outputs are medically accurate and align with established clinical guidelines. Which foundational concept addresses the trustworthiness and ethical behavior of the model in this critical application?Foundation Models
- 107.A government agency is using an AI system to process citizen requests for social benefits. They are concerned that if a request is denied, the citizen should be able to understand the specific reasons for the denial. Which Responsible AI concept is primarily addressed by ensuring that the AI's decision-making process can be clearly communicated to end-users?Responsible AI
- 108.An AI-powered content moderation system for a social media platform is being developed. The system needs to quickly identify and remove harmful content, such as hate speech or graphic violence. However, there's a risk that the AI might over-moderate, leading to censorship of legitimate content, or under-moderate, failing to remove harmful posts. To balance these trade-offs and ensure the system aligns with societal values, the development team is establishing clear policies for content classification, defining escalation paths for ambiguous cases to human moderators, and implementing regular audits by an independent ethics committee. Which Responsible AI concept is being comprehensively addressed by these actions?Responsible AI
- 109.A team is developing an AI system for predictive policing. Given the potential for significant societal impact, they are conducting a thorough assessment to identify and mitigate potential ethical, social, and legal risks *before* deployment. This assessment includes evaluating fairness, transparency, privacy implications, and potential for misuse. What is this comprehensive proactive evaluation process called?Responsible AI
- 110.A developer is iterating quickly on an application that uses a foundation model for creative writing. They frequently need to adjust the model's behavior, such as making its output more concise, more imaginative, or adhere to a specific persona, without retraining the entire model or even fine-tuning it. Which method allows for flexible and immediate control over a foundation model's output characteristics purely through its input?Foundation Models
- 111.A financial institution is developing a fraud detection system. They need to analyze vast amounts of transaction data, identify anomalous patterns that may indicate fraudulent activity, and continuously improve their detection models. Due to the sensitive nature of the data, all data processing and model training must occur within a private network and not traverse the public internet. Which combination of AWS services should they use to achieve this securely and efficiently?AWS Services for AI/ML and Generative AI
- 112.A company is mandated to ensure that all data used for training and inference of their AI/ML models remains within a specific geographical region to comply with data residency regulations. Which AWS security and compliance feature is primarily responsible for enforcing this requirement?AWS Services for AI/ML and Generative AI
- 113.A novel AI system is being developed to predict the likelihood of recidivism (reoffending) for individuals in the justice system. The system's predictions could influence sentencing, parole decisions, and resource allocation. Given the profound impact on individuals' lives and potential for societal harm, a multidisciplinary team is conducting an extensive evaluation covering ethical, legal, and social implications, including potential biases, fairness concerns, privacy risks, and accountability mechanisms, before any deployment. Which crucial Responsible AI practice is this team undertaking?Responsible AI
- 114.A company is deploying an AI system for automated hiring. To adhere to Responsible AI best practices, human recruiters review all final hiring decisions made by the AI before an offer is extended. This practice helps to mitigate potential biases or errors from the AI. What is this practice commonly known as?Responsible AI
- 115.A research team needs to train a custom deep learning model for a novel image classification task. They require a platform that provides integrated tools for data labeling, model training, hyperparameter tuning, and deployment, with full control over the underlying infrastructure and algorithms. Which AWS service offers this comprehensive, end-to-end machine learning lifecycle management?AWS Services for AI/ML and Generative AI
- 116.A data analytics team is developing a new system to automatically classify customer feedback emails into categories like 'billing inquiry', 'technical support', or 'feature request'. They are considering using a pre-trained foundation model. Which of the following best describes why a foundation model would be a suitable choice for this task, particularly if the team has limited labeled training data for their specific categories?Foundation Models
- 117.A data scientist is building a machine learning model to categorize customer feedback. They need a service that can automatically identify and extract entities like product names, organization names, and sentiment from unstructured text data without requiring extensive custom model training. Which AWS service is best suited for this task?AWS Services for AI/ML and Generative AI
- 118.A data scientist is working with a large language model (LLM) for a natural language processing task. To improve the model's performance on a specific, niche dataset without incurring the high computational cost of full fine-tuning, they decide to train only a small fraction of the model's parameters while keeping the majority of the pre-trained weights frozen. This technique is known as:Foundation Models
- 119.A credit scoring AI model is found to consistently give lower scores to individuals residing in certain postal codes, even when controlling for other financial indicators. Upon investigation, it's discovered that these postal codes historically correspond to lower-income areas, and the training data implicitly learned to associate location with creditworthiness, regardless of individual applicant merits. This leads to unfair outcomes for a specific demographic group. Which type of Responsible AI issue is this model exhibiting?Responsible AI
- 120.A data scientist is working on a project that involves generating synthetic data for training a new machine learning model. The project requires high-quality, diverse, and contextually relevant data that mimics real-world distributions. Which of the following foundation model types is BEST suited for this task?Foundation Models
- 121.A startup is building an intelligent assistant that needs to understand user queries, retrieve relevant information from a proprietary knowledge base, and then generate a coherent and accurate answer based on the retrieved information. The goal is to minimize factual inaccuracies (hallucinations) and provide up-to-date responses. Which technique is best suited for this architecture?Foundation Models
- 122.A company is developing a new customer service chatbot and needs to select a foundation model. The primary constraint is a limited budget for operational expenses, meaning the cost associated with running the model for each user interaction must be minimized. Which factor should the team prioritize when selecting the model?Foundation Models
- 123.A large e-commerce company wants to leverage generative AI to automatically create engaging product descriptions based on product attributes (e.g., color, material, features, price). They need a service that allows them to experiment with different leading foundation models, customize them with their proprietary product data, and then integrate the best-performing model into their existing product catalog management system through an API. Which AWS service is designed for this purpose?AWS Services for AI/ML and Generative AI
- 124.A security auditor is reviewing an AWS account where a data scientist is training machine learning models using Amazon SageMaker. The auditor needs to verify that all API calls made to SageMaker, including model creation, training job starts, and endpoint deployments, are logged and can be audited for compliance purposes. Which AWS service should the auditor examine to find these detailed records?AWS Services for AI/ML and Generative AI
- 125.A research team is exploring how to enable an AI model to understand the hierarchical relationships and semantic similarities between words, such as knowing that 'apple' and 'banana' are both 'fruits', and 'fruit' is a type of 'food'. Which AI/ML concept is best suited for representing these complex linguistic relationships?AI/ML and Generative AI Fundamentals
- 126.A team is training a large language model (LLM) for a specific domain, such as legal documents. They have a vast amount of general text data but only a limited, highly specialized dataset of legal texts. To adapt the LLM effectively to the legal domain, what is the most appropriate technique, assuming the base LLM is already pre-trained on general data?AI/ML and Generative AI Fundamentals
- 127.A data scientist is tasked with preparing a dataset for an AI/ML model that will predict whether a customer will renew their subscription. The dataset contains a 'Customer ID' column, which uniquely identifies each customer. Which of the following actions should the data scientist take regarding this column during data preparation?AI/ML and Generative AI Fundamentals
- 128.A data scientist is preparing a dataset of customer reviews for sentiment analysis using a machine learning model. The reviews are free-form text. To enable the model to understand the semantic relationships between words (e.g., 'good' is similar to 'excellent', 'bad' is similar to 'terrible'), which technique should be applied?AI/ML and Generative AI Fundamentals
- 129.An AI/ML team has developed a large language model (LLM) and deployed it for customer service. Users complain that the LLM sometimes generates plausible-sounding but factually incorrect responses, especially when queried about obscure topics or details not explicitly present in its training data. What is the term for this phenomenon in generative AI?AI/ML and Generative AI Fundamentals
- 130.A data engineer is preparing a dataset for an AI/ML model that will predict the probability of a customer clicking on an advertisement. The dataset contains numerical features with vastly different scales, such as 'age' (ranging from 18-90) and 'income' (ranging from $20,000-$500,000). To prevent features with larger numerical values from disproportionately influencing the model's learning process, what preprocessing technique should be applied?AI/ML and Generative AI Fundamentals
- 131.A financial institution wants to develop a generative AI model to synthesize realistic, but entirely fictional, transaction data for testing new fraud detection algorithms. The goal is to create data that mimics the statistical properties of real transactions without exposing any actual customer information. Which characteristic of generative AI is most critical for this specific use case?AI/ML and Generative AI Fundamentals
- 132.A data engineer is preparing a large dataset for training a machine learning model that will predict customer churn. The dataset contains various features, including `CustomerID`, `SubscriptionDate`, `LastLogin`, and `MonthlyCharges`. Which of these features would typically be considered a 'label' in a supervised learning context?AI/ML and Generative AI Fundamentals
- 133.A team is building a machine learning model to detect anomalies in network traffic, identifying unusual patterns that might indicate a cyberattack. They have a massive dataset of normal network traffic but very few examples of actual cyberattacks. Which type of learning approach is most suitable for this scenario, given the imbalanced and rare nature of the 'attack' class?AI/ML and Generative AI Fundamentals
- 134.A data scientist is preparing a dataset for an AI/ML model that will predict the probability of a customer clicking on an advertisement. The dataset includes features like 'Age', 'Income', and 'TimeSpentOnSite'. Which of the following data preparation steps is crucial for ensuring that features with different scales (e.g., Age in years, Income in thousands) do not disproportionately influence the model?AI/ML and Generative AI Fundamentals
- 135.A financial institution is implementing a machine learning model to detect fraudulent transactions. The model is designed to flag transactions that deviate significantly from a customer's typical spending patterns. This approach is an example of which type of machine learning task?AI/ML and Generative AI Fundamentals
- 136.A healthcare provider is using a machine learning model to predict the likelihood of a patient developing a certain disease. They discover that the model consistently underpredicts the risk for a specific demographic group, leading to missed early interventions. This situation is an example of what ethical concern in AI?AI/ML and Generative AI Fundamentals
- 137.A financial institution is implementing an AI system to detect anomalies in transaction data. The system needs to identify unusual patterns without prior examples of fraudulent activity, as new types of fraud constantly emerge. Which type of machine learning approach is best suited for this scenario?AI/ML and Generative AI Fundamentals
- 138.A data scientist is preparing a dataset for an image classification task where the goal is to identify different species of birds. The dataset contains thousands of images, each with a corresponding label indicating the bird species. Which type of machine learning task does this scenario represent?AI/ML and Generative AI Fundamentals
- 139.A team is developing a generative AI model to create new musical compositions. They want the generated music to exhibit a high degree of creativity and unexpected variations, rather than strictly adhering to common musical patterns. Which parameter should they adjust to encourage more diverse and surprising outputs?AI/ML and Generative AI Fundamentals
- 140.A data scientist is preparing a dataset for an AI/ML model that will predict house prices. The dataset contains various features, including 'number_of_bedrooms', 'square_footage', and 'zip_code'. Which of these features would most likely require one-hot encoding before being fed into a typical machine learning algorithm?AI/ML and Generative AI Fundamentals
- 141.A startup is developing an AI-powered system to generate marketing copy for various products. The system needs to create diverse and engaging text that is grammatically correct and contextually relevant. Which generative AI architecture is best known for its ability to produce coherent and high-quality sequential data, making it particularly suitable for such text generation tasks?AI/ML and Generative AI Fundamentals
- 142.An AI research team is exploring different types of neural networks for processing sequential data, such as speech recognition or predicting stock prices. They need a network architecture that can effectively remember information from previous steps in a sequence and use it to inform future predictions. Which neural network type is specifically designed for this purpose?AI/ML and Generative AI Fundamentals
- 143.A data scientist is preparing a dataset for an AI/ML model that will predict customer churn. One of the features in the dataset is 'Customer_ID', which is a unique alphanumeric identifier for each customer. What is the most appropriate action for the data scientist to take regarding this feature before training the model?AI/ML and Generative AI Fundamentals
- 144.A team is developing a large language model (LLM) for a specialized legal domain, such as patent law. They have access to a vast, general-purpose LLM pre-trained on a massive amount of diverse text data. However, they also have a smaller, highly specific dataset of legal documents and patent filings. To adapt the general LLM to excel in the specific legal domain, which technique should they employ?AI/ML and Generative AI Fundamentals
- 145.A social media platform is developing a generative AI model to create personalized short video clips for users based on trending topics. The team wants to ensure the model can understand and retain context over very long sequences of information, allowing it to generate cohesive narratives that span several minutes of video. Which key mechanism enables modern generative AI models to effectively handle these long-range dependencies?AI/ML and Generative AI Fundamentals
- 146.A financial institution is implementing a machine learning model to detect fraudulent transactions. The model is designed to classify transactions as either 'fraudulent' or 'legitimate'. Given that detecting actual fraud is extremely critical, which evaluation metric should the institution prioritize to ensure that as many fraudulent transactions as possible are caught, even if it means flagging some legitimate transactions incorrectly?AI/ML and Generative AI Fundamentals
- 147.A research team is developing a generative AI model to create novel protein structures based on a sequence of amino acids. They need an architecture that can understand long-range dependencies within the sequence and process the input in parallel for efficiency. Which neural network architecture is best suited for this task?AI/ML and Generative AI Fundamentals
- 148.A developer is building a generative AI model to create realistic images of animals. After initial training, the model produces images that are blurry and lack fine details. Which of the following generative model architectures is best suited to address this issue by improving the sharpness and quality of generated images?AI/ML and Generative AI Fundamentals
- 149.A machine learning engineer is developing a real-time anomaly detection system for network intrusion. The system needs to identify unusual patterns in network traffic without prior examples of what an 'intrusion' looks like. The engineer has access to a large dataset of normal network behavior. Which type of machine learning approach is most suitable for this task?AI/ML and Generative AI Fundamentals
- 150.An e-commerce company wants to implement a system that suggests products to customers based on their browsing history, past purchases, and items viewed by similar users. Which machine learning task is primarily responsible for powering such a system?AI/ML and Generative AI Fundamentals