Courseiva

PDE · topic practice

Storing the Data practice questions

The Storing the Data domain covers choosing and configuring Google Cloud storage and database services: Cloud SQL, Spanner, BigQuery, Cloud Storage, Bigtable, Firestore, and Memorystore. Questions test matching workload characteristics — scale, consistency, latency, schema, cost — to the right service, plus governance controls like VPC Service Controls, IAM, CMEK, and BigQuery external tables.

Courseiva uses original exam-style practice questions designed for learning and revision. The goal is to understand the concepts, recognise exam patterns, and improve through explanations — not memorise copied exam dumps.

Editorial oversight:Johnson Ajibi· MSc IT Security, IEEE Senior Member
20 questionsDomain: Storing the Data

What the exam tests

What to know about Storing the Data

Match each workload to the correct storage service using consistency, scale, latency, and schema requirements, then apply the right governance controls. The single most important thing: know when Spanner, BigQuery, Cloud Storage, Bigtable, or Firestore is the correct answer, and why the alternatives fail.

Selecting Spanner for globally consistent, horizontally scalable relational workloads with SQL joins and automatic failover.

Using BigQuery external tables or BigLake to query Parquet and other formats in Cloud Storage without loading.

Designing BigQuery schemas with nested and repeated fields to denormalize sessions and avoid joins.

Applying VPC Service Controls perimeter to restrict BigQuery and Cloud Storage access and prevent exfiltration.

Watch out for

Common Storing the Data exam traps

  • ▸Choosing Cloud SQL for global scale or Bigtable for SQL joins instead of Spanner, which supports both strong consistency and relational queries.
  • ▸Assuming BigQuery external tables perform like native tables; querying Cloud Storage directly is slower and lacks clustering and partitioning benefits.
  • ▸Treating VPC Service Controls as IAM only; it enforces network perimeters and does not replace least-privilege IAM roles.

Practice set

Storing the Data questions

20 questions · select your answer, then reveal the explanation

Question 1mediummultiple choice
Read the full Storing the Data explanation →

A data engineer is building a data lake on Google Cloud and needs to separate raw ingested data, curated/cleaned data, and processed/aggregated data. Which Cloud Storage bucket structure is recommended?

Question 2mediummultiple choice
Read the full Storing the Data explanation →

A company stores sensitive data in BigQuery and needs to encrypt certain columns with customer-managed encryption keys (CMEK) while using BigQuery's analytics capabilities. What should they do?

A data engineer needs to design a Bigtable row key for a time-series IoT application where each device sends data every second. The query pattern is to retrieve all data for a specific device over a time range. Which row key design minimizes hotspots?

A data engineer is designing a Bigtable row key for a time-series application that records temperature sensor readings every second. To avoid hotspotting, they want to distribute writes across all nodes. Which row key design is best?

A company uses Cloud Spanner for a global e-commerce platform. They have a table of orders and a table of order items. To optimize performance for queries that join these tables on order_id, which Spanner schema design feature should they use?

A company is designing a data lake on Cloud Storage with different zones. They need to enforce data retention so that objects in the 'raw' zone are automatically deleted after 1 year. Which TWO actions should they take? (Choose 2 correct options)

A global fintech company needs a database that can serve transactional (OLTP) and analytical (OLAP) workloads with strong consistency. They require high availability and PostgreSQL compatibility. Which TWO Google Cloud databases meet these requirements? (Choose 2 correct options)

A company runs a global financial application requiring strong consistency across continents with 99.999% availability. They need to store transaction data with ACID properties and sub-10ms write latency from any region. Which storage service meets all requirements?

An IoT application writes sensor readings to Cloud Bigtable with a row key of 'deviceID#timestamp'. The team notices high write latency and hotspots on a few nodes. Which row key design change would most likely improve performance?

A company wants to use BigQuery to query data stored in Cloud Storage as Parquet files without loading the data into BigQuery storage. Which feature should they use?

Question 11mediummultiple choice
Read the full Storing the Data explanation →

A healthcare company must encrypt data in BigQuery with customer-managed keys (CMEK). They want to control the key lifecycle independently. Which approach should they take?

A company uses Cloud Storage to store sensitive customer data. They need to restrict access to the data so that only requests from within a specific VPC network are allowed, and block all access from the public internet. Which TWO configurations should they implement? (Choose 2.)

Question 13mediummultiple choice
Read the full Storing the Data explanation →

An organization uses Cloud Storage to store backup files. They want to automatically delete files older than 90 days, and after deletion, move remaining files to Nearline storage if not accessed for 30 days. Which Cloud Storage feature should they configure?

A data engineer is designing a Bigtable row key for a time-series dataset where each row represents a sensor reading. The team expects high write throughput and wants to avoid hotspots. Which row key design is BEST?

A company needs a fully managed, PostgreSQL-compatible database that supports both transactional (OLTP) and analytical (OLAP) workloads with low latency. They want to minimize operational overhead. Which two Google Cloud services should they consider? (Choose two.)

A company stores data in a Cloud Storage bucket with versioning enabled. They want to automatically delete objects that are noncurrent (i.e., previous versions) after 30 days, and also delete the current version if it is older than 365 days. Which three Object Lifecycle Management conditions can be used together? (Choose three.)

Question 17mediummultiple choice
Read the full Storing the Data explanation →

A company needs to store petabytes of time-series IoT sensor data and query it with single-digit millisecond latency at millions of reads per second. The data has a simple key-value structure with timestamps. Which Google Cloud database is MOST appropriate?

A multinational corporation needs a globally distributed database that supports strong consistency, SQL queries, and automatic failover across regions. They also want to optimize join performance for parent-child relationships. Which TWO features of Cloud Spanner should they use?

Question 19mediummultiple choice
Read the full Storing the Data explanation →

A company needs to store petabytes of time-series IoT sensor data and query it with single-digit millisecond latency at millions of reads per second. The data has a simple key-value structure with timestamps. Which Google Cloud database is MOST appropriate?

You are designing a Cloud Storage bucket to hold sensitive financial documents that must not be deleted or overwritten for 7 years. After the retention period, the documents can be deleted automatically. Which configuration should you use?

Free account

Track your progress over time

Create a free account to save your results and see which topics improve across sessions.

Focused Storing the Data sessions

Start a Storing the Data only practice session

Every question in these sessions is drawn from the Storing the Data domain — nothing else.

Related practice questions

Related PDE topic practice pages

Move into related areas when this topic feels solid.

Frequently asked questions

What does the PDE exam test about Storing the Data?
Match each workload to the correct storage service using consistency, scale, latency, and schema requirements, then apply the right governance controls. The single most important thing: know when Spanner, BigQuery, Cloud Storage, Bigtable, or Firestore is the correct answer, and why the alternatives fail.
How should I use these practice questions?
Select your answer before revealing the explanation. Then read why each option is right or wrong — this active recall approach builds retention far faster than re-reading notes.
Can I practise just Storing the Data questions in a focused session?
Yes — the session launcher on this page draws every question from the Storing the Data domain. Use a 10-question session first to gauge your baseline, then move to 20 or 30 once the weak spots are clear.
Where can I practise other PDE topics?
Use the topic links above to move to related areas, or go back to the PDE question bank to see all topics.
Are these real exam questions or dumps?
These are original practice questions written to test the same concepts the PDE exam covers. They are not copied from any real exam or dump site.