1. A global manufacturing company needs to store machine telemetry data from thousands of devices. This data needs to be queried quickly for real-time dashboards and predictive maintenance. The data structure is dynamic and can change frequently as new sensors are added. Which Azure non-relational data service is the MOST appropriate for this scenario?
Describe how to work with non-relational data on Azure
A.Azure Data Lake Storage Gen2
B.Azure Blob Storage
C.Azure Table Storage
D.Azure Cosmos DB (MongoDB API)
Show answerAnswer
D. Azure Cosmos DB (MongoDB API)
Azure Cosmos DB with the MongoDB API is ideal for this scenario due to its flexible schema for dynamic data, high throughput for telemetry, and low-latency query capabilities for real-time dashboards. It natively supports document models, which align well with evolving sensor data.
2. A manufacturing company uses legacy applications that rely heavily on Apache Cassandra for their operational data store. They want to migrate these applications to Azure without significant code changes. Which Azure Cosmos DB API would provide the most straightforward migration path?
Describe how to work with non-relational data on Azure
A.SQL API
B.Cassandra API
C.Gremlin API
D.Table API
Show answerAnswer
B. Cassandra API
The Cassandra API for Azure Cosmos DB provides wire protocol compatibility with Apache Cassandra. This allows existing Cassandra applications to connect to Cosmos DB with minimal or no code changes, making it the most straightforward migration path.
3. A data analyst is querying a large table in Azure SQL Database that stores historical sales data. Queries frequently filter by a `SaleDate` column, but performance is slow. The `SaleDate` column is currently not indexed. Which basic management task should the analyst recommend to improve query performance on this column?
Describe how to work with relational data on Azure
A.Update statistics on the `SaleDate` column
B.Create a non-clustered index on the `SaleDate` column
C.Rebuild the table containing `SaleDate`
D.Change the `SaleDate` column data type to `DATE`
Show answerAnswer
B. Create a non-clustered index on the `SaleDate` column
Creating an index on a frequently filtered column like `SaleDate` significantly speeds up data retrieval by allowing the database engine to quickly locate relevant rows without scanning the entire table.
4. A web application generates a high volume of small, transient messages that need to be processed asynchronously by backend workers. The messages should be ordered, and the system needs to guarantee at-least-once delivery. Which Azure non-relational data service is best suited for this messaging pattern?
Describe how to work with non-relational data on Azure
A.Azure Data Lake Storage Gen2
B.Azure Queue Storage
C.Azure Cosmos DB
D.Azure Blob Storage
Show answerAnswer
B. Azure Queue Storage
Azure Queue Storage is specifically designed for storing large numbers of messages that can be retrieved by an application. It provides reliable, at-least-once delivery and supports message ordering (though not strictly guaranteed for all scenarios, it's generally maintained for simple queues). Its cost-effectiveness and simplicity make it ideal for decoupling application components with asynchronous messaging.
5. A research institution stores petabytes of raw scientific data, including high-resolution images, video files, and experimental results. This data is rarely accessed after initial ingestion but must be retained for regulatory compliance and potential future analysis for decades. Which Azure Blob Storage access tier is the most cost-effective for this scenario?
Describe how to work with non-relational data on Azure
A.Cool
B.Archive
C.Hot
D.Premium
Show answerAnswer
B. Archive
The Archive tier is the most cost-effective for rarely accessed, long-term data storage. While it has higher retrieval costs and latency, its storage cost is significantly lower, making it ideal for archival purposes.
6. A financial company uses Azure Database for PostgreSQL to store customer transaction data. Due to increasing demand during peak hours, the database is experiencing performance bottlenecks, specifically with read-heavy analytical queries. The company needs to offload these read workloads from the primary database to improve overall responsiveness. Which Azure Database for PostgreSQL feature should be utilized?
Describe how to work with relational data on Azure
A.Point-in-time Restore
B.High Availability
C.Connection Pooling
D.Read Replicas
Show answerAnswer
D. Read Replicas
Read Replicas in Azure Database for PostgreSQL are designed to offload read-heavy workloads from the primary server, improving performance and scalability for read operations.
7. A data engineer is working with an Azure SQL Database. They need to create a table that will store employee records, and each employee must have a unique employee ID. This ID should be automatically generated by the database when a new employee record is inserted. Which column property should be used for the Employee ID column?
Describe how to work with relational data on Azure
A.DEFAULT
B.PRIMARY KEY
C.NOT NULL
D.IDENTITY
Show answerAnswer
D. IDENTITY
The IDENTITY property (or AUTO_INCREMENT in other SQL dialects) in Azure SQL Database automatically generates unique, sequential numbers for a column when new rows are inserted. This perfectly matches the requirement for an automatically generated, unique employee ID.
8. A company is migrating an on-premises SQL Server database to Azure. The database contains several features, including SQL Server Agent jobs, cross-database queries, and Distributed Transaction Coordinator (DTC). The company needs a fully managed service that minimizes refactoring efforts while providing nearly 100% compatibility with their existing SQL Server environment. Which Azure relational data service should they choose?
Describe how to work with relational data on Azure
A.Azure Database for MySQL
B.Azure SQL Database Managed Instance
C.SQL Server on Azure Virtual Machines
D.Azure SQL Database
Show answerAnswer
B. Azure SQL Database Managed Instance
Azure SQL Database Managed Instance offers near 100% compatibility with the latest SQL Server (Enterprise Edition) database engine. It supports features like SQL Server Agent, cross-database queries, and DTC, which are not available in Azure SQL Database (PaaS) but are crucial for minimizing refactoring efforts during migration from an on-premises SQL Server.
9. A research institution is collecting massive amounts of scientific data, including high-resolution images, video streams, and experimental results, totaling hundreds of terabytes. This data needs to be stored cost-effectively for long-term retention (over 10 years) and accessed very rarely, typically only for specific historical research projects. When accessed, retrieval time can be several hours. Which Azure storage solution is the most appropriate for this scenario?
Describe how to work with non-relational data on Azure
A.Azure Files Premium
B.Azure Blob Storage (Archive tier)
C.Azure Blob Storage (Cool tier)
D.Azure Data Lake Storage Gen2
Show answerAnswer
B. Azure Blob Storage (Archive tier)
The Archive tier of Azure Blob Storage is specifically designed for very rarely accessed, long-term archival data. It offers the lowest storage costs but has the highest retrieval costs and latency, which aligns perfectly with the requirement for cost-effective long-term retention and acceptance of retrieval times of several hours.
10. A global e-commerce platform uses Azure Cosmos DB to store product catalog information. To ensure high availability and low-latency reads for users worldwide, the platform needs to distribute its data across multiple Azure regions and allow any region to accept write operations. Which Cosmos DB feature should be enabled?
Describe how to work with non-relational data on Azure
A.Point-in-time restore
B.Change Feed
C.Session Consistency
D.Multi-region writes
Show answerAnswer
D. Multi-region writes
Multi-region writes in Azure Cosmos DB allow all configured regions to accept write operations, providing globally distributed write capabilities and enhancing both availability and write latency for geographically dispersed users.
11. A data engineer is working with an Azure SQL Database. The database stores sensor readings where each reading has a unique identifier that needs to be automatically generated and incremented for every new record. Which property should be assigned to the SensorID column to achieve this?
Describe how to work with relational data on Azure
A.UNIQUE
B.PRIMARY KEY
C.IDENTITY
D.DEFAULT
Show answerAnswer
C. IDENTITY
The IDENTITY property automatically generates sequential numeric values for a column, which is precisely what's needed for an auto-incrementing unique identifier.
12. A manufacturing company uses legacy applications that rely heavily on Apache Cassandra for their wide-column data model. The company plans to migrate these applications to Azure without rewriting their data access layer. They need a fully managed, scalable NoSQL database service that is compatible with existing Cassandra drivers and tools. Which Azure Cosmos DB API should they choose?
Describe how to work with non-relational data on Azure
A.Azure Cosmos DB for Cassandra API
B.Azure Cosmos DB for Gremlin (Graph) API
C.Azure Cosmos DB for NoSQL (Core) API
D.Azure Cosmos DB for MongoDB API
Show answerAnswer
A. Azure Cosmos DB for Cassandra API
The Azure Cosmos DB for Cassandra API provides wire protocol compatibility with Apache Cassandra. This allows existing Cassandra applications to connect to Cosmos DB without code changes, leveraging existing drivers and tools, which is crucial for a migration without rewriting the data access layer.
13. A small startup is developing a mobile game that needs to store player high scores and basic player profiles (e.g., player ID, username, last login). The data model is simple, primarily key-value pairs, and the startup needs a highly scalable, low-cost solution without complex indexing or advanced querying capabilities. Which Azure non-relational data service is the most suitable?
Describe how to work with non-relational data on Azure
A.Azure Table Storage
B.Azure Blob Storage
C.Azure Data Lake Storage Gen2
D.Azure Cosmos DB (SQL API)
Show answerAnswer
A. Azure Table Storage
Azure Table Storage is a low-cost, highly scalable NoSQL key-value store. It's ideal for simple data models like high scores and basic profiles that require fast lookups without the overhead of complex database features.
14. A data architect is designing a relational database for an e-commerce platform in Azure. The database includes a 'Categories' table and a 'Products' table, where each product belongs to a category. When a category is deleted, all associated products should also be automatically deleted to maintain data consistency. Which referential integrity action should be configured on the foreign key constraint?
Describe how to work with relational data on Azure
A.SET NULL
B.RESTRICT
C.NO ACTION
D.CASCADE
Show answerAnswer
D. CASCADE
The CASCADE referential integrity action ensures that when a row in the parent table (Categories) is deleted, all corresponding rows in the child table (Products) are also automatically deleted, maintaining consistency.
15. A global online streaming service needs to store user watch history, recommendations, and preferences. This data needs to be available with single-digit millisecond latency worldwide and support multiple data models (document, graph). The service anticipates massive, unpredictable spikes in traffic. Which Azure service is designed to meet these requirements?
Describe how to work with non-relational data on Azure
A.Azure SQL Database
B.Azure Data Lake Storage Gen2
C.Azure Cosmos DB
D.Azure Database for PostgreSQL
Show answerAnswer
C. Azure Cosmos DB
Azure Cosmos DB is a globally distributed, multi-model database service that guarantees single-digit millisecond latency, elastic scalability, and supports various APIs including document (SQL, MongoDB) and graph (Gremlin), perfectly aligning with the requirements for global availability, low latency, and varied data models.
16. A media company needs to store millions of video and audio files, each potentially gigabytes in size. These files will be ingested, processed, and then served to users via a streaming application. The storage solution must be highly durable, scalable to petabytes, and accessible via HTTP/HTTPS. Which Azure non-relational data service is best suited for storing these large media files?
Describe how to work with non-relational data on Azure
A.Azure Table Storage
B.Azure SQL Database
C.Azure Cosmos DB
D.Azure Blob Storage
Show answerAnswer
D. Azure Blob Storage
Azure Blob Storage is Microsoft's object storage solution for the cloud, designed to store massive amounts of unstructured data, such as video and audio files. It offers high durability, petabyte-scale scalability, and direct access via HTTP/HTTPS, making it ideal for media streaming applications.
17. A media company is developing a new video streaming platform. They need to store millions of video and audio files, each potentially gigabytes in size. These files will be frequently accessed by users worldwide, requiring high availability and low-latency streaming. Which Azure Storage service is the most appropriate for this requirement?
Describe how to work with non-relational data on Azure
A.Azure Files
B.Azure Blob Storage
C.Azure Table Storage
D.Azure Data Lake Storage Gen2
Show answerAnswer
B. Azure Blob Storage
Azure Blob Storage is designed for storing large amounts of unstructured object data, such as video and audio files. Its ability to handle large files, provide high availability, and integrate with CDNs makes it ideal for streaming platforms.
18. A data analytics team needs to store petabytes of raw sensor data from IoT devices for long-term analysis. The data will be ingested in real-time, often consisting of small, individual files. The team requires a cost-effective storage solution that supports a hierarchical namespace for organization and is optimized for analytical workloads, including integration with Spark and Hadoop. Which Azure service is the most appropriate choice?
Describe how to work with non-relational data on Azure
A.Azure Blob Storage (Standard)
B.Azure Table Storage
C.Azure Data Lake Storage Gen2
D.Azure Cosmos DB
Show answerAnswer
C. Azure Data Lake Storage Gen2
Azure Data Lake Storage Gen2 is built on Azure Blob Storage but adds a hierarchical namespace and is optimized for big data analytics workloads. It provides file system semantics, file-level security, and is highly scalable and cost-effective for petabytes of data, making it ideal for IoT sensor data analysis with Spark and Hadoop integration.
19. A company is migrating an on-premises SQL Server database to Azure. The database contains sensitive customer information and requires high availability with automated failover capabilities. The company needs to minimize management overhead for patching and backups while retaining full SQL Server engine compatibility. Which Azure relational data service should they choose?
Describe how to work with relational data on Azure
A.Azure SQL Database single database
B.Azure SQL Database Hyperscale
C.Azure SQL Managed Instance
D.SQL Server on Azure Virtual Machines
Show answerAnswer
C. Azure SQL Managed Instance
Azure SQL Managed Instance offers near 100% compatibility with the latest SQL Server (Enterprise Edition) database engine, combined with high availability and automated management, making it ideal for lift-and-shift migrations with minimal application changes.
20. A data analyst is querying an Azure SQL Database that contains customer order data. They need to retrieve all orders placed in the last 30 days. Which SQL clause is most appropriate for filtering the results based on the order date?
Describe how to work with relational data on Azure
A.HAVING
B.ORDER BY
C.GROUP BY
D.WHERE
Show answerAnswer
D. WHERE
The WHERE clause is used to filter rows based on a specified condition before any grouping or ordering takes place.
21. A financial institution needs to store audit logs for regulatory compliance. These logs are generated continuously and need to be retained for seven years. While the logs are rarely accessed after initial ingestion, they must be highly durable and accessible when needed. Cost-effectiveness for long-term storage is a primary concern. Which Azure Blob Storage access tier is the most appropriate for these audit logs?
Describe how to work with non-relational data on Azure
A.Archive
B.Cool
C.Premium
D.Hot
Show answerAnswer
A. Archive
The Archive tier is the most cost-effective solution for storing data that is rarely accessed but needs to be retained for long periods, like audit logs for regulatory compliance. While retrieval latency is higher (hours), the significantly lower storage costs over seven years make it the ideal choice when access frequency is low and cost is paramount.
22. A data engineering team needs to store petabytes of raw sensor data from IoT devices for long-term analysis. The data will be ingested continuously and processed using big data analytics frameworks like Apache Spark. A hierarchical namespace is required for efficient organization and access control. Which Azure non-relational data service is most suitable?
Describe how to work with non-relational data on Azure
A.Azure Queue Storage
B.Azure Cosmos DB
C.Azure Data Lake Storage Gen2
D.Azure Table Storage
Show answerAnswer
C. Azure Data Lake Storage Gen2
Azure Data Lake Storage Gen2 is built on Azure Blob Storage and provides a hierarchical file system, making it ideal for big data analytics workloads. It supports petabyte-scale data, continuous ingestion, and integrates well with frameworks like Apache Spark.
23. A company is migrating its legacy on-premises SQL Server databases to Azure. The databases contain complex stored procedures, triggers, and SQL Agent jobs that are critical for business operations. The company needs full SQL Server engine compatibility and administrative control. Which Azure relational data service is the most suitable choice?
Describe how to work with relational data on Azure
A.Azure SQL Managed Instance (PaaS)
B.Azure SQL Database (PaaS)
C.Azure Synapse Analytics Dedicated SQL Pool (PaaS)
D.Azure Database for MySQL (PaaS)
Show answerAnswer
A. Azure SQL Managed Instance (PaaS)
Azure SQL Managed Instance offers near 100% compatibility with the latest SQL Server (on-premises) database engine, providing full SQL Server features, administrative control, and lift-and-shift capability for existing applications.
24. A global online gaming platform uses Azure Cosmos DB to store player profiles, game statistics, and leaderboards. They need to ensure that regardless of where a player connects from, they experience very low latency when accessing their data. The platform also requires high availability, even in the event of a regional outage. Which Azure Cosmos DB feature should be implemented to meet these requirements?
Describe how to work with non-relational data on Azure
A.Change Feed
B.Stored Procedures
C.Indexing Policies
D.Global Distribution
Show answerAnswer
D. Global Distribution
Azure Cosmos DB's Global Distribution feature allows data to be replicated across multiple Azure regions. This provides low-latency access to data for users worldwide by serving data from the closest region, and ensures high availability through automatic failover in case of a regional outage.
25. A retail company uses Azure Blob Storage to store millions of customer receipts as PDF files. Due to compliance regulations, these receipts must be retained for a minimum of seven years. However, after the first year, they are rarely accessed. To optimize storage costs while ensuring compliance, which Azure Blob Storage access tier should be used for receipts older than one year?
Describe how to work with non-relational data on Azure
A.Archive
B.Hot
C.Premium
D.Cool
Show answerAnswer
A. Archive
The Archive access tier is the most cost-effective for rarely accessed data with flexible latency requirements. It's suitable for long-term retention of data that can tolerate several hours of retrieval time, which aligns with the scenario of receipts rarely accessed after one year but needing to be retained for seven.
Azure Cosmos DB's MongoDB API allows developers to interact with Cosmos DB using MongoDB client drivers and tools, providing a globally distributed, multi-model database service with the flexibility of the MongoDB document model.
Compatible with existing MongoDB applications and tools.
Offers flexible schema for diverse and evolving data.
Provides global distribution and high availability.
An API for Azure Cosmos DB that provides wire protocol compatibility with Apache Cassandra, enabling easy migration of existing Cassandra applications.
Supports Cassandra Query Language (CQL).
Leverages Cosmos DB's global distribution and scalability.
Ideal for migrating existing Cassandra workloads to Azure with minimal code changes.
A database object that provides fast lookup of data in a table, similar to an index in a book. It speeds up data retrieval operations on a table at the cost of additional storage and slower data modification operations.
Improves query performance, especially for WHERE clauses and JOINs.
Can be clustered (sorts and stores data rows) or non-clustered (separate structure).
Should be created on frequently queried or joined columns.
A column property in SQL Server (and Azure SQL DB) that automatically generates unique, sequential numeric values for a column when new rows are inserted.
A fully managed cloud database service that provides near 100% compatibility with the latest SQL Server database engine, suitable for lift-and-shift migrations.
Near 100% SQL Server compatibility
Supports SQL Server Agent, cross-database queries, DTC
The CASCADE referential action automatically deletes or updates dependent rows in the child table when the corresponding parent row is deleted or updated.
Maintains data consistency.
Automates deletion/update of related records.
Can be powerful but requires careful design to avoid unintended data loss.
Azure Blob Storage is Microsoft's object storage solution for the cloud. It is optimized for storing massive amounts of unstructured data, such as text or binary data.
Stores unstructured data (blobs) like images, videos, audio, documents.
Highly scalable to petabytes of data.
Accessible via HTTP/HTTPS from anywhere in the world.
A fully managed, highly compatible SQL Server database engine service in Azure, designed for lift-and-shift migrations of on-premises SQL Server workloads to the cloud.
Near 100% SQL Server compatibility.
Automated backups, patching, and high availability.
Offers a private network environment for enhanced security.
A core capability of Azure Cosmos DB that allows your data to be replicated across any number of Azure regions worldwide, providing low-latency access and high availability.
Replicates data across multiple Azure regions.
Ensures low-latency data access for users globally.
Provides high availability with automatic failover during regional outages.
Azure Role-Based Access Control (RBAC) for Storage
Flip card
A system that provides fine-grained access management to Azure resources, including Storage accounts, by assigning roles to Azure Active Directory identities (users, groups, service principals).
Integrates directly with Azure Active Directory (AAD).
Grants permissions based on predefined or custom roles.
A deployment option for Azure Database for PostgreSQL that automatically scales compute capacity based on demand and pauses billing during periods of inactivity, optimized for intermittent and unpredictable workloads.
A feature in Azure SQL Database that encrypts the entire database, backups, and transaction log files at rest, providing encryption without application changes.
Encrypts data at rest (database files, backups, logs).
Transparent to applications (no code changes needed).
Managed by the Azure SQL Database service by default.
A LEFT JOIN (or LEFT OUTER JOIN) returns all records from the left table, and the matching records from the right table. If there is no match, NULLs are returned for the right side.
Questions are original practice items written to match the published exam objectives. Step2Study is not affiliated with or endorsed by any certification body.