AIP-C01 Operational Efficiency And Optimization Practice Question
Which THREE metrics are best used for setting up auto-scaling policies on a SageMaker endpoint?
Answer choices
Why each option matters
Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.
Correct answer & explanation
✓
InvocationsPerInstance
InvocationsPerInstance, CPUUtilization, and GPUUtilization are the standard metrics to trigger scaling actions.
Answer analysis
Option-by-option breakdown
For each option: why learners choose it and why it is or isn't the right answer here.
- ✗
IAMUserLoginCount
Why it's wrong here
This is irrelevant to endpoint load.
- ✓
InvocationsPerInstance
Why this is correct
A direct measure of traffic load per instance.
- ✓
CPUUtilization
Why this is correct
An indicator of compute-bound processing load.
- ✗
S3RequestLatency
Why it's wrong here
This is for storage access, not inference.
- ✓
GPUUtilization
Why this is correct
An indicator of model inference load on GPUs.
About these practice questions
One of 181 original AIP-C01 practice questions on Courseiva, each with a full explanation and wrong-answer analysis — not exam dumps or protected exam content. Learn why practice questions differ from exam dumps →
JA
Written and reviewed by Johnson Ajibi, MSc IT Security
Senior Network & Security Engineer · founder of Courseiva
Last reviewed August 2026 · checked against the official Amazon Web Services exam blueprint
This AIP-C01 practice question is part of Courseiva's free Amazon Web Services certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the AIP-C01 exam.