Courseiva
Data Operations and Support →mediumMultiple Choice

DEA-C01 Data Operations and Support Practice Question

A data engineer is using AWS Step Functions to orchestrate a complex ETL workflow that includes multiple AWS Glue jobs, Amazon EMR steps, and AWS Lambda functions. The engineer notices that on rare occasions, the entire workflow fails due to a transient error in one of the Lambda functions. The engineer wants to make the workflow more resilient without changing the overall architecture. Which approach is the MOST effective?

⚠ Common exam trap

The trap here is thinking that increasing Lambda resources or enabling tracing will resolve transient errors, when the real solution is to implement automated retries.

Answer choices

Why each option matters

Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.

Correct answer & explanation

✓

Add a Retry state in the Step Functions state machine for the Lambda task with exponential backoff and a maximum number of attempts.

Adding a Retry state in Step Functions with exponential backoff and a maximum attempts limit is the most effective way to handle transient errors in Lambda functions. It automatically retries the failed task, increasing the likelihood of success without manual intervention, and is a standard resilience pattern in Step Functions.

Answer analysis

Option-by-option breakdown

For each option: why learners choose it and why it is or isn't the right answer here.

  • ✗

    Enable AWS X-Ray tracing for the Lambda function to identify the root cause of the errors.

    Why it's wrong here

    X-Ray tracing provides visibility into the Lambda function's execution and can help diagnose issues, but it does not automatically recover from transient errors. While useful for debugging, it does not make the workflow more resilient. The engineer needs an automated recovery mechanism, such as retries, to handle transient failures.

  • ✗

    Modify the Lambda function code to catch exceptions and return a success status to avoid workflow failure.

    Why it's wrong here

    Catching exceptions and returning success would mask the failure and potentially lead to incorrect data processing. It does not actually resolve the transient error; it simply hides it. This approach could result in data inconsistencies and is not a recommended practice for building resilient workflows. Retrying the operation is a better solution.

  • ✓

    Add a Retry state in the Step Functions state machine for the Lambda task with exponential backoff and a maximum number of attempts.

    Why this is correct

    Adding a Retry state in Step Functions allows the workflow to automatically retry the Lambda function when it fails due to transient errors. Configuring exponential backoff and a maximum attempt count ensures that temporary issues are handled gracefully without manual intervention. This directly improves the workflow's resilience and is a best practice for handling transient failures.

  • ✗

    Configure the Lambda function to have a longer timeout and increase its memory size.

    Why it's wrong here

    Increasing timeout and memory can help if the Lambda function is timing out or running out of memory, but it does not address transient errors such as network issues or service throttling. These errors are often temporary and can be resolved by retrying the operation. Therefore, this approach may not effectively improve resilience against transient failures.

Quick reference

Cloud Service Model Comparison

ModelYou ManageProvider ManagesExamples
IaaSOS, runtime, apps, dataHardware, hypervisor, networkingEC2, Azure VMs, GCP Compute Engine
PaaSApps and dataOS, runtime, middleware, hardwareElastic Beanstalk, Azure App Service
SaaSData and settings onlyEverything elseMicrosoft 365, Salesforce, Workday
FaaS / ServerlessFunction code onlyInfra, scaling, runtimeLambda, Azure Functions, Cloud Run
CaaSContainers and appsKubernetes, OS, hardwareEKS, AKS, GKE

About these practice questions

Courseiva writes every DEA-C01 question from scratch — 1,321 in total, each with an explanation and a wrong-answer breakdown. None are copied from real exams or dumps. Learn why practice questions differ from exam dumps →

How Courseiva writes practice questions · Editorial policy

JA

Written and reviewed by Johnson Ajibi, MSc IT Security

Senior Network & Security Engineer · founder of Courseiva

Last reviewed September 2026 · checked against the official Amazon Web Services exam blueprint

This DEA-C01 practice question is part of Courseiva's free Amazon Web Services certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the DEA-C01 exam.