Courseiva

DEA-C01 Data Ingestion and Transformation Practice Question

A data engineer needs to transform JSON data into CSV format using AWS Glue. The transformation is simple and must be executed on a schedule. Which Glue component is MOST suitable?

⚠ Common exam trap

DEA-C01 often tests the misconception that a Crawler performs transformation, when in fact Crawlers only catalog schema and ETL jobs do the actual data processing.

Answer choices

Why each option matters

Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.

Correct answer & explanation

✓

Glue ETL job

A Glue ETL job is the component designed to run transformation scripts (Spark or Python shell) on data, and it can be scheduled via triggers. For a simple JSON-to-CSV transformation on a schedule, an ETL job is the correct choice.

Answer analysis

Option-by-option breakdown

For each option: why learners choose it and why it is or isn't the right answer here.

  • ✗

    Glue Crawler

    Why it's wrong here

    A Glue Crawler infers schemas and populates the Data Catalog; it performs no JSON-to-CSV conversion and cannot be scheduled as an ETL job. It is tempting because crawlers run on schedules and are the usual first step before transformation, and would be correct when the goal is classifying raw S3 data into catalog tables.

  • ✗

    Glue Data Catalog

    Why it's wrong here

    Glue Data Catalog stores table metadata and schemas; it executes no transformation logic and cannot run on a schedule. It is tempting because the JSON source must be catalogued before ETL reads it, and it would be correct when the requirement is centralised metadata for Athena, Redshift, or crawler-discovered schemas.

  • ✗

    Glue Development Endpoint

    Why it's wrong here

    A Glue Development Endpoint provisions an interactive notebook environment for authoring and debugging ETL scripts; it does not itself schedule or run production jobs. It is tempting during iterative development of custom PySpark transforms, and would be correct when an engineer needs REPL-style testing against sample data before deployment.

  • ✓

    Glue ETL job

    Why this is correct

    A Glue ETL job runs Apache Spark transformations, converting JSON to CSV, and supports scheduling via triggers. Glue crawlers only catalogue data and DataBrew is interactive, so the ETL job matches the simple scheduled transformation requirement.

About these practice questions

One of 1,321 original DEA-C01 practice questions on Courseiva, each with a full explanation and wrong-answer analysis — not exam dumps or protected exam content. Learn why practice questions differ from exam dumps →

How Courseiva writes practice questions · Editorial policy

JA

Written and reviewed by Johnson Ajibi, MSc IT Security

Senior Network & Security Engineer · founder of Courseiva

Last reviewed September 2026 · checked against the official Amazon Web Services exam blueprint

This DEA-C01 practice question is part of Courseiva's free Amazon Web Services certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the DEA-C01 exam.