Courseiva

MLA-C01 Deployment and Orchestration of ML Workflows Practice Question

A company wants to deploy a trained XGBoost model for batch inference on a large dataset stored in S3. The inference job should be cost-effective and does not require real-time responses. Which SageMaker inference option should they use?

⚠ Common exam trap

MLA-C01 often tests the cost/latency trade-off between Batch Transform and Asynchronous Inference — candidates pick Asynchronous because it sounds 'batch-like,' but Asynchronous Inference is for near-real-time large-payload requests, not bulk offline scoring.

Answer choices

Why each option matters

Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.

Correct answer & explanation

✓

SageMaker Batch Transform

SageMaker Batch Transform is designed for offline, high-throughput inference on large datasets stored in S3, with no persistent endpoint and no real-time requirement. It is the most cost-effective option for this scenario because you pay only for the duration of the batch job.

Answer analysis

Option-by-option breakdown

For each option: why learners choose it and why it is or isn't the right answer here.

  • ✓

    SageMaker Batch Transform

    Why this is correct

    SageMaker Batch Transform runs inference over large S3 datasets on managed instances that terminate when the job completes, avoiding the cost of a persistently running endpoint. This satisfies the stem's cost-effectiveness requirement and its lack of any real-time response need.

  • ✗

    SageMaker real-time endpoint

    Why it's wrong here

    A real-time endpoint holds compute continuously and returns synchronous responses, so it bills for idle capacity while processing a large S3 dataset that needs no immediate answers. Batch Transform reads from S3 and shuts down after the job. Real-time endpoints suit low-latency interactive predictions.

  • ✗

    SageMaker Asynchronous Inference

    Why it's wrong here

    Asynchronous Inference queues requests and returns results via S3, but it targets large payloads or long processing times needing near-real-time retrieval, not whole-dataset scoring. Batch Transform handles the full S3 dataset cost-effectively. Asynchronous suits requests exceeding synchronous payload limits.

  • ✗

    SageMaker Serverless Inference

    Why it's wrong here

    Serverless Inference scales to zero and suits intermittent, unpredictable traffic with latency tolerance, but it caps payload size and duration, making it unsuitable for scoring a large S3 dataset. Batch Transform processes the entire dataset in one job. Serverless suits spiky, low-volume request patterns.

Quick reference

AWS S3 Storage Class Comparison

Storage ClassMin DurationRetrievalUse Case
S3 StandardNoneImmediateFrequently accessed data
S3 Standard-IA30 daysImmediateInfrequent access, rapid retrieval
S3 One Zone-IA30 daysImmediateNon-critical infrequent data
S3 Intelligent-TieringNoneImmediate–hoursUnknown or changing access patterns
S3 Glacier Instant90 daysMillisecondsArchive with instant retrieval
S3 Glacier Flexible90 daysMinutes–hoursArchive, flexible retrieval
S3 Glacier Deep Archive180 daysHoursLong-term compliance archive

About these practice questions

This MLA-C01 question is part of Courseiva's 665-question bank — original exam-style content with full explanations and wrong-answer analysis, never real exam questions or exam dumps. Learn why practice questions differ from exam dumps →

How Courseiva writes practice questions · Editorial policy

JA

Written and reviewed by Johnson Ajibi, MSc IT Security

Senior Network & Security Engineer · founder of Courseiva

Last reviewed September 2026 · checked against the official Amazon Web Services exam blueprint

This MLA-C01 practice question is part of Courseiva's free Amazon Web Services certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the MLA-C01 exam.