Microsoft Azure Data Fundamentals practice questions
248 free questions with answers and explanations.
- 101.A development team is building a new application that processes a large volume of small, independent messages. Each message represents a task that needs to be processed asynchronously by a worker service. The system must ensure that each message is processed at least once, even if a worker fails, and messages should be automatically removed after successful processing. Which Azure Storage service is designed for this type of asynchronous message queuing?Describe how to work with non-relational data on Azure
- 102.A global e-commerce platform uses Azure Cosmos DB to store product catalog information. The platform needs to serve product details to customers with consistently low latency, regardless of their geographical location. To achieve this, the product catalog data must be available in multiple Azure regions, and the application must be able to write to any of these regions. Which Cosmos DB feature ensures this capability?Describe how to work with non-relational data on Azure
- 103.An application uses Azure Database for PostgreSQL and experiences inconsistent query performance, particularly during peak hours. The database administrator suspects that some queries are not efficiently utilizing available indexes. Which basic management task should the administrator perform to analyze and potentially resolve this issue?Describe how to work with relational data on Azure
- 104.A data engineer is designing a data warehouse solution in Azure Synapse Analytics. They have a very large fact table, `FactSales`, that will be frequently queried with filters on a `RegionID` column. The goal is to distribute the data to minimize data movement during queries that filter by `RegionID`. Which table distribution method should be chosen for `FactSales`?Describe how to work with relational data on Azure
- 105.A healthcare provider is building an application to manage patient records. These records must be consistent across all replicas, even in the event of network partitions, to ensure data integrity and prevent conflicting updates. Which Azure Cosmos DB consistency level should be chosen?Describe how to work with non-relational data on Azure
- 106.A global e-commerce platform uses Azure Cosmos DB to store product catalog information. To ensure maximum uptime and data durability, the platform requires that data be replicated across multiple Azure regions. Additionally, in the event of an entire region failure, the application must automatically failover to a healthy region without manual intervention. Which Cosmos DB feature primarily fulfills these requirements?Describe how to work with non-relational data on Azure
- 107.A company is designing a new database in Azure SQL Database to store highly sensitive customer credit card information. They need to ensure that specific columns containing this data are encrypted within the database and that only authorized applications or users can decrypt and access the plaintext data. This encryption should prevent even database administrators from seeing the sensitive data. Which encryption technology should be chosen?Describe how to work with relational data on Azure
- 108.A development team is building an application that will use Azure Database for PostgreSQL. They need to ensure that the database can handle sudden, unpredictable spikes in traffic without manual intervention and only pay for the compute resources consumed. Which compute tier should they select for the PostgreSQL server?Describe how to work with relational data on Azure
- 109.A startup is building a new social media platform where user connections (friends, followers) are highly interconnected and need to be queried efficiently for recommendations and network analysis. Which Azure Cosmos DB API is most appropriate for modeling and querying this type of data?Describe how to work with non-relational data on Azure
- 110.A startup is building a social media application where users can follow each other, and posts can have many comments and likes. They need a database that can efficiently store and query complex relationships between users, posts, comments, and likes. Which Azure Cosmos DB API is the most suitable for this use case?Describe how to work with non-relational data on Azure
- 111.A security auditor needs to ensure that access to sensitive documents stored in Azure Blob Storage is strictly controlled based on user roles and identities managed in Azure Active Directory (AAD). Which Azure storage management feature should be implemented to achieve this?Describe how to work with non-relational data on Azure
- 112.A company is storing a large number of images and videos in Azure Blob Storage. They want to ensure that these media files are accessible only to authorized users and applications, and that access is granted based on the principle of least privilege. Which security mechanism should be used to provide granular, role-based access control (RBAC) to specific containers and blobs?Describe how to work with non-relational data on Azure
- 113.A global gaming company needs to store player profiles, game statistics, and leaderboards. This data is semi-structured, accessed frequently by millions of players worldwide, and requires extremely low latency (single-digit milliseconds). The solution must also be able to scale elastically to handle peak loads. Which Azure non-relational data service is the most appropriate choice?Describe how to work with non-relational data on Azure
- 114.A developer is creating an application that requires a simple key-value store to maintain user preferences and configuration settings. Each item will have a unique identifier (key) and associated attributes (values). The schema for these attributes might vary slightly between different types of items but is generally flat. High throughput and low latency are important, but global distribution is not a primary concern. Which Azure non-relational data service is the most cost-effective and suitable choice?Describe how to work with non-relational data on Azure
- 115.A global e-commerce platform uses Azure Cosmos DB to store product catalog information. To ensure high availability and low latency for customers worldwide, the platform needs to allow writes to occur in multiple Azure regions simultaneously. Which Azure Cosmos DB feature enables this capability?Describe how to work with non-relational data on Azure
- 116.A startup is building a new social media platform where users can follow each other, and posts can have nested comments and likes. The application requires storing complex relationships between entities (users, posts, comments) and performing efficient traversal queries. Which Azure Cosmos DB API is most appropriate for this data model?Describe how to work with non-relational data on Azure
- 117.A data administrator is managing an Azure SQL Database that is experiencing periodic slowdowns in query execution, even for queries that historically performed well. The underlying data distribution in some tables has changed significantly due to frequent inserts and updates. Which maintenance task should the administrator perform to address this issue?Describe how to work with relational data on Azure
- 118.A data engineer is designing a solution to store millions of small files (averaging 10 KB each) from IoT devices. These files are generated continuously and need to be stored cost-effectively for potential future analysis. The files are rarely accessed after initial ingestion, and retrieval performance is not critical, with latencies of several minutes or even hours being acceptable. Which Azure Blob Storage access tier is the most appropriate for this scenario?Describe how to work with non-relational data on Azure
- 119.A media company is developing a new video streaming platform. They need to store millions of video files, each potentially gigabytes in size. These files will be accessed frequently by users worldwide. Which Azure non-relational data service is most suitable for storing these video assets?Describe how to work with non-relational data on Azure
- 120.A global e-commerce company uses Azure Database for MySQL to power its storefront. To improve read scalability and reduce latency for users in different geographical regions, the company wants to distribute read-only copies of its database. Which feature of Azure Database for MySQL should be configured?Describe how to work with relational data on Azure
- 121.An analytics team needs to store raw sensor data from IoT devices for long-term analysis. The data arrives in varying formats and volumes, and the team intends to use various big data processing frameworks like Apache Spark and Hadoop for analysis. Which Azure non-relational data service is specifically designed to support these requirements for large-scale analytics?Describe how to work with non-relational data on Azure
- 122.A data engineer is working with an Azure SQL Database. They need to retrieve all customer records along with any orders they have placed. If a customer has not placed any orders, their information should still be included in the result set, with NULLs for the order-related columns. Which type of JOIN should the engineer use?Describe how to work with relational data on Azure
- 123.A startup is building a new mobile application that needs to store semi-structured data for user profiles, product catalogs, and in-app purchases. The data schema is expected to evolve frequently, requiring flexibility. The application needs to scale globally and deliver low-latency responses. Which type of non-relational database is best suited for this evolving, semi-structured data?Describe how to work with non-relational data on Azure
- 124.A global online gaming platform uses Azure Cosmos DB to store player profiles, game statistics, and leaderboards. To ensure the best possible performance and lowest latency for players worldwide, the platform needs to allow data writes from any region where players are active. Which Azure Cosmos DB feature enables this capability?Describe how to work with non-relational data on Azure
- 125.A web application generates a high volume of small, transient messages that need to be processed asynchronously by a backend service. The messages must be durable and guaranteed to be delivered at least once, even if the backend service is temporarily unavailable. Which Azure non-relational data service is most suitable for this messaging scenario?Describe how to work with non-relational data on Azure
- 126.A data analytics team needs to store petabytes of raw sensor data from IoT devices for long-term retention and future big data processing using tools like Apache Spark. The solution must support a hierarchical namespace for efficient organization and access control. Which Azure non-relational data service is best suited for this purpose?Describe how to work with non-relational data on Azure
- 127.A research institution collects massive amounts of scientific data, including high-resolution images, video files, and experimental results. This data needs to be stored for decades for future analysis and compliance. Access to older data is infrequent but must be possible. The primary concern is minimizing storage costs while ensuring long-term data integrity. Which Azure storage solution is most appropriate?Describe how to work with non-relational data on Azure
- 128.A data analytics team needs to store petabytes of raw sensor data from IoT devices for long-term analysis. The data will be ingested continuously and processed in large batches using Apache Spark. The solution requires a hierarchical namespace for efficient data organization and security, and compatibility with the Hadoop ecosystem. Which Azure non-relational data service is most appropriate for this scenario?Describe how to work with non-relational data on Azure
- 129.A data architect is designing a data warehouse solution using Azure Synapse Analytics. They are creating a large fact table that will store billions of rows of IoT sensor data. This data will primarily be queried based on a `SensorID` column, which has a high number of distinct values and is frequently used in JOIN and WHERE clauses. The goal is to optimize query performance for these specific types of queries. Which table distribution strategy should the architect choose for this fact table in a Synapse Dedicated SQL Pool?Describe how to work with relational data on Azure
- 130.A developer is building a new application that processes a large volume of small, independent messages. Each message needs to be processed asynchronously and reliably, even if the processing service is temporarily unavailable. The messages should be automatically removed after a configurable time period if not processed. Which Azure non-relational data service is best suited for this scenario?Describe how to work with non-relational data on Azure
- 131.A data analytics team needs to store petabytes of raw sensor data from IoT devices for long-term analysis. The data arrives in various formats (CSV, JSON, Avro) and will be processed by big data analytics engines like Apache Spark. The storage solution must support hierarchical namespaces, fine-grained access control, and be optimized for analytical workloads. Which Azure Storage service is the most suitable?Describe how to work with non-relational data on Azure
- 132.A developer is creating an application that requires a simple key-value store to maintain configuration settings and user session data. The application needs to handle a high volume of transactions with consistent low latency. Which Azure non-relational data service should the developer choose?Describe how to work with non-relational data on Azure
- 133.A company is developing an IoT solution that collects telemetry data from thousands of devices. Each device sends a small message (e.g., sensor readings, status updates) every few seconds. These messages need to be temporarily stored in a buffer before being processed by a backend service. The solution requires high throughput for message ingestion and guaranteed delivery to the processing service. Which Azure non-relational service is most suitable for this message queuing scenario?Describe how to work with non-relational data on Azure
- 134.A data engineer is designing a data warehouse solution using Azure Synapse Analytics. They have a large fact table that stores transactional data. This table will be frequently queried with filters on specific columns (e.g., 'CustomerID', 'ProductID') and will be involved in many JOIN operations. To optimize query performance, which table distribution strategy should be chosen for this fact table in a dedicated SQL pool?Describe how to work with relational data on Azure
- 135.A database administrator is managing an Azure SQL Database. The database experiences periods of high query load, leading to some queries taking longer than expected. The DBA suspects that the query optimizer is making suboptimal choices for certain complex queries. To help the query optimizer make better decisions and improve query performance without rewriting the queries, which database management task should the DBA perform regularly?Describe how to work with relational data on Azure
- 136.A data team is working with an Azure SQL Database. They need to retrieve all product names and their associated category names. Some products might not have a category assigned, and these products should still appear in the result set with a NULL value for the category name. Which type of JOIN should be used?Describe how to work with relational data on Azure
- 137.A financial institution needs to store audit logs for regulatory compliance. These logs are generated continuously, must be retained for seven years, and are accessed very infrequently, typically only for audits. When accessed, retrieval time can be several hours. Which Azure Blob Storage access tier is the most cost-effective for this scenario?Describe how to work with non-relational data on Azure
- 138.A security auditor needs to ensure that access to sensitive documents stored in Azure Blob Storage is granted only to specific users and groups within Azure Active Directory (AAD). The auditor requires a solution that provides fine-grained control over individual blobs and containers, leveraging existing AAD identities. Which mechanism should be used?Describe how to work with non-relational data on Azure
- 139.A company is developing a social media application where users can follow each other, and posts can have many comments. They need a database that can efficiently store and query highly connected data, such as finding all friends of a friend, or all comments on a post and their authors. Which Azure Cosmos DB API is best suited for modeling and querying this type of relationship-heavy data?Describe how to work with non-relational data on Azure
- 140.A development team is building a new application that uses Azure SQL Database. The application will store monetary values, such as product prices and transaction amounts. These values require exact precision and scale, and should not be subject to floating-point inaccuracies. Which data type should the team use to store these monetary values?Describe how to work with relational data on Azure
- 141.A developer is creating a relational database in Azure SQL Database. They need to define a column that stores monetary values and ensures precise calculations without loss of precision, which is critical for financial transactions. Which data type should they choose for this column?Describe how to work with relational data on Azure
- 142.A large e-commerce platform uses Azure Cosmos DB to store its product catalog. The platform operates globally, with users accessing the catalog from various continents. To ensure the fastest possible response times for all users, the company wants to configure Cosmos DB so that data is written to the nearest regional replica and immediately available for reads from any replica, even if it means a slight chance of reading data that hasn't fully propagated globally. Which Azure Cosmos DB consistency level should the company choose?Describe how to work with non-relational data on Azure
- 143.A gaming company is developing a new multiplayer online game that needs to store player profiles, game statistics, and leaderboards. The data requires extremely low latency reads and writes, global distribution to serve players worldwide, and the ability to scale elastically to handle millions of concurrent users. Which Azure non-relational database service is best suited for these requirements?Describe how to work with non-relational data on Azure
- 144.A startup is developing a new social media platform where users can follow each other, post updates, and form groups. The application needs to efficiently query complex relationships between users, posts, and groups, such as 'find all friends of friends who liked a specific post' or 'identify all users in a group who are also following a particular influencer'. Which Azure non-relational data service is best suited for this type of highly connected data and complex relationship querying?Describe how to work with non-relational data on Azure
- 145.A data engineering team is designing a data ingestion pipeline for real-time telemetry data from millions of IoT devices. The data arrives as JSON documents and needs to be stored in a NoSQL database that offers flexible schema and can handle high throughput with low latency. The team also requires the ability to query this data using a familiar SQL-like language. Which Azure Cosmos DB API is the MOST appropriate choice?Describe how to work with non-relational data on Azure
- 146.A global logistics company needs to store tracking information for millions of packages. Each package's tracking status changes frequently, and the data for each package (e.g., current location, delivery attempts, customs information) can have a highly variable and evolving structure. The system requires high throughput for updates and low latency for retrieving a package's current status. Which data model is MOST appropriate for this scenario?Describe core data concepts
- 147.A healthcare provider needs to store patient medical records, which include scanned documents, X-ray images, and video recordings of consultations. These files are typically large and need to be accessed securely and efficiently, but do not require complex querying of their internal content. Which Azure storage option is MOST appropriate for this scenario?Describe core data concepts
- 148.A manufacturing company uses IoT sensors to monitor the temperature and pressure of its machinery. The sensors generate millions of data points per second. The company needs to store this data efficiently for real-time dashboards and predictive maintenance, with a focus on fast ingestion and time-based queries. Which type of database is best suited for this specific workload?Describe core data concepts
- 149.A data scientist is working with a dataset that contains information about customer purchases. Each record includes 'CustomerID', 'ProductID', 'Quantity', and 'OrderDate'. To analyze the total quantity of products sold per customer over time, the data scientist needs to group the data by 'CustomerID' and 'OrderDate' and then sum 'Quantity'. This operation is an example of which data processing concept?Describe core data concepts
- 150.A data engineering team is responsible for transforming raw data from various sources into a clean, structured format suitable for analytical reporting. This process involves extracting data from source systems, cleaning and normalizing it, and then loading it into a data warehouse. Which common data workload best describes this entire sequence of operations?Describe core data concepts