Practice GCP-ADP Data Preparation And Ingestion questions with full explanations on every answer.
Start practicing
Data Preparation And Ingestion — choose a session length
Free · No account required
Click any question to see the full explanation and answer options, or start a focused practice session above.
You need to store sensitive data in GCS. Which feature ensures data is encrypted at rest?
2Which GCS storage class is most cost-effective for data accessed only once per year?
3You are designing a streaming pipeline in Dataflow. You need to ensure that data is processed in the order it was generated, even if it arrives late. What do you use?
4You need to load CSV files from GCS to BigQuery. The schema changes frequently. Which approach handles schema evolution automatically?
5You need to perform data cleansing on a large dataset in Dataprep. Which execution engine should you choose for scalability?
6You need to ingest data from an external API that requires an OAuth token. Which tool is best for executing this periodically?
7You are processing streaming data in Dataflow and notice 'stuck' elements causing pipeline latency. Which feature helps debug this?
8You want to run a Dataflow pipeline on a schedule. What should you use?
9You need to set up an alert when a GCS bucket exceeds a certain size. Which tool do you use?
10You need to transfer files from an Amazon S3 bucket to GCS. What is the most efficient method?
11You are migrating a high-throughput on-premises database to BigQuery. You need to ensure zero downtime. What is your best strategy?
12You need to provide temporary access to a GCS file for a third party. What should you use?
13You need to ingest IoT data into BigQuery. You need to handle messages arriving in millions per second. Which service acts as the buffer?
14You are running a Dataflow job and notice that it is consuming too much memory and crashing. How do you optimize this?
15You need to track who accessed which GCS bucket. Which service provides this audit trail?
16You are processing streaming data in Pub/Sub and need to archive every message into GCS without writing custom code. What should you use?
17You need to ensure that deleted files in GCS can be restored for 30 days. What should you configure?
18You are cleaning data using Dataprep. You want to save the final dataset in BigQuery. What do you do?
19You are using Dataflow to read from Pub/Sub. The pipeline is failing due to malformed messages. How can you handle these without crashing?
20What is the primary benefit of using partitioned tables in BigQuery?
21You need to ingest log data into BigQuery from GCS files that arrive sporadically. Which service is best to trigger this?
22You are running a Dataflow job that joins two large datasets. Which join strategy should you avoid to prevent OOM errors?
23Which GCP tool provides a managed way to run Apache Airflow workflows?
24You need to anonymize PII (Personally Identifiable Information) before loading data into BigQuery. Which tool is best?
25You need to ingest data from a legacy database into Google Cloud. The database has strict network egress controls. How should you connect?
26You need to delete all files in a GCS bucket older than 90 days. What is the most efficient approach?
27You are using Dataflow with autoscaling. What metric determines whether the worker pool scales up or down?
28You need to verify the integrity of a large file uploaded to GCS. What is the standard way to do this?
29You need to perform a complex windowing operation on streaming data in Dataflow. Which windowing strategy is best for session-based activity?
30You are using BigQuery and need to load data from an external file that is not in GCS. Which method is recommended?
31Which service should you use to ingest streaming data from mobile devices?
32You are using Dataflow with a custom container. Which command do you use to specify the container image?
33You are using Dataprep to prepare data. How can you share your work with a team?
34How do you restrict access to a specific GCS bucket to only members of a specific team?
35Which tool is used to monitor the performance of your Dataflow pipelines?
36You are using BigQuery and need to query a large dataset that is partitioned by day. Which clause should you always include for efficiency?
37You need to ingest small files into Cloud Storage using a command-line tool. Which command is most appropriate?
38In Dataflow, what is the impact of using a high number of workers for a small dataset?
39You need to perform a rolling update on a Dataflow job without losing current state. How do you do this?
40Which TWO of the following are valid ways to ingest data into BigQuery?
41Which TWO factors should you consider when choosing a GCS storage class?
42Which THREE of the following are Dataflow windowing types?
43Which TWO of the following are benefits of using Cloud Storage?
44Which THREE features does Dataprep provide?
45Which TWO of the following are common Dataflow pipeline patterns?
46Which TWO of the following are GCP ingestion tools?
47Which THREE of the following are valid GCS storage classes?
48Which THREE steps are involved in an effective Dataflow pipeline development?
49Which TWO are common causes of Dataflow pipeline failures?
50Which THREE are features of Pub/Sub?
51Which TWO of the following are valid ways to monitor Dataflow performance?
52Which TWO are methods to secure GCS data?
53Which THREE of the following are valid BigQuery table types?
54Which TWO of the following are key benefits of using Dataflow?
55Which TWO are valid methods to trigger a Dataflow job?
56Which THREE of the following are necessary for a production-ready Dataflow pipeline?
57You need to ingest large amounts of unstructured data into Cloud Storage from an on-premises data center with limited bandwidth. Which service should you choose to ensure the most cost-effective and secure transfer?
58Your team uses Dataprep by Trifacta to clean data before loading it into BigQuery. You notice that the column header names contain inconsistent casing and special characters. Which Dataprep transformation should you use to standardize these headers globally?
59You are building a streaming pipeline using Dataflow to process sensor data arriving via Pub/Sub. You need to handle out-of-order data by allowing events to arrive late. Which Dataflow concept must you configure?
60You are troubleshooting a Dataflow job that is running slower than expected when writing data to BigQuery. You suspect the issue is related to hot keys. What is the recommended strategy to mitigate this?
61Your organization requires that all data ingested into Cloud Storage be encrypted at rest using keys managed by you, not Google. Which feature should you implement?
62You have a large CSV file in Cloud Storage that needs to be loaded into BigQuery. The file contains a nested JSON structure in one column. How should you best prepare this data?
63You are configuring a Cloud Data Fusion pipeline to ingest data from an external SQL database. You need to ensure that only rows modified within the last hour are ingested. Which feature should you use?
64Which TWO of the following are recommended practices when designing a Dataflow pipeline for high-throughput streaming ingestion?
65You are optimizing a Dataflow job. Which THREE of the following actions can help improve job performance and reduce costs?
66You are planning to ingest log data into Google Cloud. Which THREE of the following services support streaming ingestion?
The Data Preparation And Ingestion domain covers the key concepts tested in this area of the GCP-ADP exam blueprint published by Google Cloud. Courseiva provides free domain-focused practice, mock exams, missed-question review, and readiness tracking across all GCP-ADP domains — no account required.
The Courseiva GCP-ADP question bank contains 66 questions in the Data Preparation And Ingestion domain. Click any question to see the full explanation and answer breakdown.
Start with a 10-question focused session to identify your baseline accuracy in this domain. Read every explanation — even for questions you answer correctly — to understand the reasoning. Once you score consistently above 80%, move to a 20–30 question session to confirm depth before moving to the next domain.
Yes — the session launcher on this page draws questions exclusively from the Data Preparation And Ingestion domain. Choose 10, 20, 30, or 50 questions for a focused session, or click individual questions to review them one by one.
Save your results, see per-domain analytics, and get readiness scores — free, for every certification.
Sign Up FreeFree forever · Every certification included