AWS Certified Data Engineer – AssociateData Storage and ManagementEasy
A data engineering team is designing a highly available and durable storage solution for frequently accessed, structured data that requires ACID transaction support. The data will be used by multiple applications for real-time analytics and operational reporting, and the team needs to scale read replicas independently. Which AWS service is best suited for this requirement?
- AAmazon RDS for PostgreSQL
- BAmazon Aurora
- CAmazon S3 Glacier Deep Archive
- DAmazon DynamoDB
Show answer & explanationAnswer & explanation
Correct answer: B. Amazon Aurora
Amazon Aurora is a MySQL and PostgreSQL-compatible relational database built for the cloud, combining the performance and availability of traditional enterprise databases with the simplicity and cost-effectiveness of open-source databases. It offers high availability, ACID compliance, and the ability to scale read replicas independently, making it ideal for the described use case.
Why the other options are wrong
- A. Amazon RDS for PostgreSQL provides a managed relational database but Aurora offers superior performance, scalability, and availability features, especially for independent read replica scaling.
- C. Amazon S3 Glacier Deep Archive is for long-term, infrequently accessed archival storage, not for frequently accessed, real-time analytics.
- D. Amazon DynamoDB is a NoSQL database suitable for key-value and document workloads, but it does not inherently provide ACID transactions in the same way a relational database does for complex structured data.
Amazon Aurora
A MySQL and PostgreSQL-compatible relational database built for the cloud, combining the performance and availability of traditional enterprise databases with the simplicity and cost-effectiveness of open-source databases.
- High performance and availability
- ACID transaction support
- Scales read replicas independently
- Cost-effective with pay-as-you-go pricing
Memory trick: Aurora shines brightest for flexible, fast, and ACID-compliant relational data.