AWS Certified Solutions Architect – ProfessionalContinuously Improve Existing SolutionsMedium
A financial institution uses an on-premises data warehouse that struggles to scale with increasing data volumes and query complexity. Data analysts frequently experience long query times, impacting reporting and decision-making. The institution wants to migrate its data warehouse to AWS to improve query performance, scalability, and reduce operational overhead. The data volume is in the petabytes, and analytical queries are complex, often involving large joins across many tables. Which AWS service is best suited for this requirement?
- AAmazon Aurora
- BAmazon Redshift
- CAmazon DynamoDB
- DAmazon Neptune
Show answer & explanationAnswer & explanation
Correct answer: B. Amazon Redshift
Amazon Redshift is a fully managed, petabyte-scale data warehouse service optimized for analytical workloads. Its columnar storage, massively parallel processing (MPP) architecture, and advanced query optimizer are specifically designed to handle complex queries over large datasets, making it ideal for this scenario.
Why the other options are wrong
- A. Amazon Aurora is a relational database optimized for OLTP (online transaction processing) workloads, not for large-scale analytical data warehousing.
- C. DynamoDB is a NoSQL database, excellent for high-performance key-value and document workloads, but not designed for complex analytical queries across petabytes of data.
- D. Amazon Neptune is a graph database, suitable for highly connected data, but not for traditional relational data warehousing and complex analytical joins.
Amazon Redshift
A fully managed, petabyte-scale data warehouse service in the cloud, optimized for fast analytical queries over large datasets.
- Columnar storage and Massively Parallel Processing (MPP) for performance.
- Scales to petabytes of data.
- Cost-effective for analytical workloads.
Memory trick: Redshift shifts your data into high gear for analytics, even at petabyte scale.