Microsoft Azure Data FundamentalsDescribe how to work with non-relational data on AzureHard

A research institution is collecting massive amounts of scientific data, including high-resolution images, video streams, and experimental results, totaling hundreds of terabytes. This data needs to be stored cost-effectively for long-term retention (over 10 years) and accessed very rarely, typically only for specific historical research projects. When accessed, retrieval time can be several hours. Which Azure storage solution is the most appropriate for this scenario?

  1. AAzure Files Premium
  2. BAzure Blob Storage (Archive tier)
  3. CAzure Blob Storage (Cool tier)
  4. DAzure Data Lake Storage Gen2
Show answer & explanation

Correct answer: B. Azure Blob Storage (Archive tier)

The Archive tier of Azure Blob Storage is specifically designed for very rarely accessed, long-term archival data. It offers the lowest storage costs but has the highest retrieval costs and latency, which aligns perfectly with the requirement for cost-effective long-term retention and acceptance of retrieval times of several hours.

Why the other options are wrong

  • A. Azure Files Premium is for high-performance file shares, which is not suitable for cost-effective, long-term, rarely accessed archival of massive unstructured data.
  • C. Cool tier is for infrequently accessed data, not rarely accessed with multi-hour retrieval tolerance; it would be more expensive for long-term archival.
  • D. Azure Data Lake Storage Gen2 is optimized for big data analytics and frequent access, not for very rare, long-term archival with high latency tolerance.

Azure Blob Storage Archive Tier

The lowest-cost storage tier for Azure Blob Storage, designed for data that is rarely accessed (typically less than once a year) and can tolerate retrieval latencies of several hours.

  • Lowest storage costs among all Blob Storage tiers.
  • Highest data retrieval costs and latency (up to 15 hours).
  • Ideal for long-term backups, compliance data, and archival.
  • Data must be rehydrated to Hot or Cool tier before it can be read.

Memory trick: For data that's 'archived' like old scrolls, the Archive tier offers a deep freeze.

More Describe how to work with non-relational data on Azure questions