Courseiva
Data Store Management →easyMultiple Choice

DEA-C01 Data Store Management Practice Question

A data engineer is building a data lake on Amazon S3 and needs to catalog metadata for a large number of CSV files stored in a folder structure. The engineer wants to use AWS Glue crawlers to automatically infer schemas and create tables in the AWS Glue Data Catalog. The crawler should run daily to detect new files and schema changes. Which configuration should the engineer use for the crawler?

⚠ Common exam trap

The trap here is choosing to create a separate table for each file, which seems granular but leads to catalog sprawl and is not how crawlers are typically used for partitioned data.

Answer choices

Why each option matters

Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.

Correct answer & explanation

✓

Set the crawler's data source to the specific S3 folder containing the CSV files, configure it to create a single schema for each S3 path, and set a daily schedule.

The crawler should be pointed to the specific S3 folder to avoid scanning irrelevant data. Configuring it to create a single schema for each S3 path groups files with the same schema into one table, which is ideal for a data lake with partitioned folders. A daily schedule ensures the catalog stays up to date with new files and schema changes.

Answer analysis

Option-by-option breakdown

For each option: why learners choose it and why it is or isn't the right answer here.

  • ✗

    Set the crawler's data source to the specific S3 folder and disable the crawler schedule, running it manually when needed.

    Why it's wrong here

    Disabling the schedule means the crawler will not automatically detect new files or schema changes, requiring manual intervention. The requirement is to run daily, so a schedule must be enabled. Manual runs are not suitable for ongoing data ingestion.

  • ✓

    Set the crawler's data source to the specific S3 folder containing the CSV files, configure it to create a single schema for each S3 path, and set a daily schedule.

    Why this is correct

    Pointing the crawler to the specific folder limits its scope to relevant data. Configuring it to create a single schema for each S3 path groups files with the same schema into one table, which is efficient for partitioned data. A daily schedule ensures new files and schema changes are detected automatically.

  • ✗

    Set the crawler's data source to the specific S3 folder and configure it to create a separate table for each file, with a daily schedule.

    Why it's wrong here

    Creating a separate table for each file would result in a large number of tables, making the Data Catalog difficult to manage and query. This is inefficient for a data lake with many files. The crawler should group files with the same schema into a single table.

  • ✗

    Set the crawler's data source to the S3 bucket root and enable 'Update the table definition in the Data Catalog' with a schedule of daily.

    Why it's wrong here

    Using the bucket root may cause the crawler to scan unrelated data and create unnecessary tables. It is better to point to the specific folder. Although the update option and schedule are correct, the data source scope is too broad, leading to inefficiency and potential catalog clutter.

Quick reference

AWS S3 Storage Class Comparison

Storage ClassMin DurationRetrievalUse Case
S3 StandardNoneImmediateFrequently accessed data
S3 Standard-IA30 daysImmediateInfrequent access, rapid retrieval
S3 One Zone-IA30 daysImmediateNon-critical infrequent data
S3 Intelligent-TieringNoneImmediate–hoursUnknown or changing access patterns
S3 Glacier Instant90 daysMillisecondsArchive with instant retrieval
S3 Glacier Flexible90 daysMinutes–hoursArchive, flexible retrieval
S3 Glacier Deep Archive180 daysHoursLong-term compliance archive

About these practice questions

One of 1,321 original DEA-C01 practice questions on Courseiva, each with a full explanation and wrong-answer analysis — not exam dumps or protected exam content. Learn why practice questions differ from exam dumps →

How Courseiva writes practice questions · Editorial policy

JA

Written and reviewed by Johnson Ajibi, MSc IT Security

Senior Network & Security Engineer · founder of Courseiva

Last reviewed September 2026 · checked against the official Amazon Web Services exam blueprint

This DEA-C01 practice question is part of Courseiva's free Amazon Web Services certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the DEA-C01 exam.