Be able to run and interpret COPY INTO, read COPY_HISTORY output to explain load state, and choose the right stage, file format, and integration for a source. The most important thing is understanding when files are loaded, skipped, or purged.
Start practicing
Data Loading, Unloading, and Connectivity — choose a session length
Free · No account required
Domain overview
This domain covers moving data into and out of Snowflake and connecting external systems. It is tested through scenario questions on COPY INTO, stages, file formats, Snowpipe, external tables, and COPY_HISTORY output, requiring you to interpret load behavior, error states, and configuration options rather than just recall syntax.
Exam objectives
COPY INTO options including PURGE, ON_ERROR, VALIDATION_MODE, and FORCE behavior on staged files
Stage types: internal named stages, external stages over S3, Azure Blob, and GCS with storage integrations
Snowpipe continuous loading with notification integrations, pipe status, and COPY_HISTORY load metadata
External tables, file formats (JSON, CSV, Parquet), and connectivity via drivers and partner tools
Assuming PURGE=TRUE deletes the stage files immediately after COPY INTO; it removes staged files only after a successful load.
Reading COPY_HISTORY rows as all failures when status and error columns show partial loads or already-loaded files being skipped.
Confusing Snowpipe auto-ingest with scheduled loads, and expecting pipes to reload files already recorded in load history.
Click any question to see the full explanation and answer options, or start a focused practice session above.
Which type of Snowflake stage is automatically created for every user and cannot be dropped or altered?
2A developer is running a COPY INTO command to load 1,000 CSV files. The requirement is that if even one row in any file fails due to a data type mismatch, the entire load operation must stop and no data should be committed to the table. Which ON_ERROR setting is required?
3A Python developer wants to upload a Pandas DataFrame to a Snowflake table as efficiently as possible without manually writing files to a local stage. Which function from the Snowflake Connector for Python should be used?
4For optimal parallel loading performance using a Snowflake virtual warehouse, what is the generally recommended compressed file size range for data files in a stage?
5A company is using an external S3 stage to load data daily. They want Snowflake to automatically delete the source files from the S3 bucket only after they have been successfully loaded into the table. Which COPY INTO option should they enable?
6When loading semi-structured data like Parquet into a Snowflake table, what is a primary advantage of using Parquet over CSV for the ingestion process?
7Which TWO statements are true regarding the behavior and management of External Tables in Snowflake?
8Refer to the exhibit. A user executes a COPY INTO command and then queries the COPY_HISTORY. Based on the output shown, what most likely happened during the load and what is the current state of the data in the SALES_DATA table?
9Refer to the exhibit. When the COPY INTO command is executed, which field delimiter will Snowflake use to parse the files located in the 'data/' folder of the stage?
10A data engineer needs to verify the structure and content of several CSV files staged in an internal Snowflake stage without actually loading the data into the production table or incurring significant compute costs. Which parameter should be used with the COPY INTO command to achieve this specific goal?
11Which Snowflake command is used to upload data files from a local file system to an internal stage?
12Which property of the FILE_FORMAT object should be adjusted if a CSV file uses a semicolon (;) instead of a comma to separate values?
13What is the purpose of the PURGE = TRUE option in a COPY INTO command?
14Which Snowflake feature allows for secure, direct communication between a customer's virtual private cloud (VPC) and the Snowflake service without using the public internet?
15A data engineer is designing a pipeline for a high-frequency stream of small JSON files arriving every minute in an S3 bucket. Why would Snowpipe be preferred over a scheduled COPY INTO command running on a dedicated virtual warehouse?
16Which Snowflake feature allows a user to download data from a Snowflake table into a local folder on their computer using the SnowSQL command-line interface?
17When using the Snowflake Connector for Python, which method is most efficient for uploading and loading large local CSV files into a Snowflake table?
18A user wants to check for potential errors in a set of staged files without actually loading the data or consuming significant warehouse credits. Which approach should they use?
19A data engineer is loading a batch of semi-structured JSON files from an external stage into a VARIANT column. They want the COPY INTO command to skip any file that contains malformed JSON without failing the entire load. Which COPY INTO option should they configure?
20A data engineer is configuring a Snowflake external stage that points to an Amazon S3 bucket. The bucket is in the same region as the Snowflake account. The engineer wants to avoid embedding long-lived AWS credentials in the stage definition and instead use a secure, temporary credential mechanism. Which authentication method should be used for the external stage?
21A data engineer is using the Snowpipe REST API to ingest data from an external stage. They need to ensure that the pipe does not reprocess files that have already been loaded. Which mechanism does Snowpipe use to track which files have been processed?
22A data engineer is loading a 4 GB CSV file from an external stage into a Snowflake table using COPY INTO. The file is compressed with gzip and has a header row. The engineer notices the load is taking longer than expected. Which action is MOST likely to improve performance?
23A data engineer needs to unload data from a Snowflake table to an external stage that references an Amazon S3 bucket. The engineer wants to ensure that the unloaded files are encrypted using a customer-managed key in AWS KMS. Which COPY INTO <location> parameter should be used to specify the KMS key?
24A data engineer is configuring a Snowpipe to automatically ingest files as they arrive in an external stage. They need to set up event notifications from the cloud provider to Snowflake. Which two components are required to enable this automated ingestion? (Choose two.)
25A data engineer has set up a Snowpipe that continuously loads JSON files from an external stage into a table. The stage references an Amazon S3 bucket with a notification integration. After several days, the engineer notices that new files are not being ingested, even though they exist in the S3 bucket. The Snowpipe is in a RUNNING state and the notification integration is active. What is the MOST likely cause of the missing data?
26A data engineer has a CSV file on their local machine and wants to load it into a Snowflake table. They do not have access to an external cloud storage bucket and want the simplest path that does not require creating a named internal stage. Which command should they use?
27A data engineer needs to load a 4.2 GB uncompressed CSV file from an internal stage into a Snowflake table. The file cannot be split because the CSV has embedded newlines within quoted fields. The engineer wants to maximize load performance. What should the engineer do?
28A data analyst needs to unload the results of a query from a Snowflake table to a local machine. The analyst wants to use the Snowflake web interface (Snowsight) to download the data as a CSV file. Which of the following is the correct approach?
29A data engineer is loading data from a local file system into a Snowflake table using the PUT command to an internal stage, followed by COPY INTO. The engineer notices that some rows are rejected due to data type mismatches. The engineer wants to capture the rejected records and continue loading valid rows. Which COPY INTO option should be used to achieve this?
30A data engineer loads JSON files from an external Azure stage into a table with a single VARIANT column. The JSON documents are newline-delimited, and each line is a separate object. Which file format type and option should be specified to correctly parse one JSON object per line?
31A user needs to load data from a CSV file stored in an external stage into a Snowflake table. The CSV file has a header row and uses a pipe (|) as the field delimiter. The user wants to ensure the header row is skipped and the pipe delimiter is recognized. Which FILE_FORMAT options should be specified in the COPY INTO command?
32A data analyst wants to query data stored in an external stage (Amazon S3) without loading it into a Snowflake table. The analyst creates an external table pointing to the stage. Which statement accurately describes how the data is accessed?
33A data engineer is configuring Snowpipe to automatically ingest files as they arrive in an external S3 stage. They must ensure the pipe loads only files matching a specific path prefix and that duplicate notifications for the same file do not cause duplicate rows. Which two configurations should they apply? (Choose two.)
34A data engineer is configuring a Snowflake storage integration to allow Snowflake to access an external S3 bucket. The engineer needs to ensure that the integration has the necessary permissions to read and write data. Which two actions must the engineer perform? (Choose two.)
35A data engineer is loading JSON data from an external stage into a Snowflake table using COPY INTO with a JSON file format. The JSON records contain nested arrays and objects. The engineer wants to load specific elements into separate columns. Which approach should the engineer use?
36A data engineer is configuring a Snowpipe to automatically load data from an external stage (Google Cloud Storage) into a Snowflake table. The engineer wants to minimize latency between file arrival and data availability. Which configuration should the engineer use?
37A data engineer is loading data from a set of CSV files stored in an external stage into a Snowflake table. The files have a header row, and the engineer wants to skip the header during loading. The engineer also wants to ensure that any rows with missing values in a NOT NULL column are skipped and logged. Which FILE_FORMAT option should be used to skip the header row?
38A data engineer is using the COPY INTO command to unload data from a Snowflake table to an external stage (Amazon S3). The engineer wants to ensure the unloaded files are encrypted and can be decrypted by the target system. Which TWO statements are true regarding the encryption of unloaded files? (Choose two.)
39A data engineer is using the COPY INTO command to load data from an external stage into a Snowflake table. The engineer wants to ensure that the load operation does not fail if some files in the stage have already been loaded previously. The engineer also wants to avoid reloading files that have already been processed. Which COPY INTO option should be used to achieve this?
40A data engineer must continuously load Parquet files arriving in an external Azure stage into a Snowflake table. The files have no consistent naming pattern and arrive at unpredictable intervals. The engineer wants Snowflake to detect and load new files automatically without running COPY on a schedule. Which Snowflake feature should be configured?
41A data engineer is troubleshooting a COPY INTO command that loads JSON files from a named external stage and is seeing unexpected NULL values in several VARIANT columns. Which TWO actions should the engineer take to diagnose how the JSON is being parsed? (Choose two.)
42A data engineer needs to load a CSV file into a Snowflake table. The file has a header row, fields are separated by commas, and string values are enclosed in double quotes. The engineer wants to ensure the header row is skipped and the string values are loaded without quotes. Which FILE_FORMAT option should be used to achieve this?
Be able to run and interpret COPY INTO, read COPY_HISTORY output to explain load state, and choose the right stage, file format, and integration for a source. The most important thing is understanding when files are loaded, skipped, or purged.
The Courseiva COF-C03 question bank contains 42 questions in the Data Loading, Unloading, and Connectivity domain. Click any question to see the full explanation and answer breakdown.
Start with a 10-question focused session to identify your baseline accuracy in this domain. Read every explanation — even for questions you answer correctly — to understand the reasoning. Once you score consistently above 80%, move to a 20–30 question session to confirm depth before moving to the next domain.
Yes — the session launcher on this page draws questions exclusively from the Data Loading, Unloading, and Connectivity domain. Choose 10, 20, 30, or 50 questions for a focused session, or click individual questions to review them one by one.
Save your results, see per-domain analytics, and get readiness scores — free, for every certification.
Sign Up FreeFree forever · Every certification included