easyMultiple Choice
PMLE Practice Question: Refer to the exhibit
Exhibit
apiVersion: serving.knative.dev/v1
kind: Service
metadata:
name: model-serving
spec:
template:
spec:
containers:
- image: gcr.io/my-project/model:v2
resources:
limits:
cpu: '2'
memory: 8Gi
startupProbe:
tcpSocket:
port: 8080
initialDelaySeconds: 60
periodSeconds: 10
containerConcurrency: 80Refer to the exhibit. A team deploys a model using Cloud Run. They notice that after scaling up, the new instances take about 90 seconds to become ready and serve requests. They want to reduce this startup time. Which configuration change is most likely to help?
⚠ Common exam trap
PMLE often tests whether candidates confuse probe timing adjustments with actual startup performance improvements, leading them to pick the startupProbe change instead of reducing image size.
Answer choices
Why each option matters
Answer the question above first, then reveal the full breakdown to understand why each option is right or wrong.
Correct answer & explanation
✓
Change the container image to use a smaller base image
Changing the container image to use a smaller base image reduces the image size and the number of layers that must be pulled and unpacked, which directly reduces container startup time on new Cloud Run instances. A smaller image means less data to download and less filesystem work before the application can start, addressing the 90-second startup delay. This is the most likely configuration change to improve cold start performance.
Answer analysis
Option-by-option breakdown
For each option: why learners choose it and why it is or isn't the right answer here.
- ✗
Reduce the startupProbe initialDelaySeconds to 30
Why it's wrong here
initialDelaySeconds delays when the startup probe first runs; it postpones readiness detection rather than shortening actual initialisation. Startup probes exist to tolerate slow-starting containers before liveness checks begin. Cutting the delay risks premature failure, and the 90-second load time remains unchanged.
- ✓
Change the container image to use a smaller base image
Why this is correct
A smaller base image reduces the bytes pulled and unpacked at instance start, shortening container initialisation before the model loads. Since the 90-second delay occurs during scale-up readiness, shrinking the image directly targets that startup time.
- ✗
Reduce the memory limit to 4Gi
Why it's wrong here
Lowering the memory limit constrains available RAM, which can slow or crash model loading rather than accelerate readiness. Memory limits exist to cap resource consumption and cost. Reducing startup time requires keeping warm instances or boosting CPU during startup, not shrinking memory.
- ✗
Increase the containerConcurrency to 100
Why it's wrong here
containerConcurrency caps simultaneous requests per instance; it does not shorten the 90-second model load and initialisation. Raising it suits high-throughput, low-latency services wanting fewer instances. Startup latency is addressed by minimum instances, startup CPU boost, or a smaller image.
Go deeper
Related to this question
About these practice questions
Courseiva writes every PMLE question from scratch — 775 in total, each with an explanation and a wrong-answer breakdown. None are copied from real exams or dumps. Learn why practice questions differ from exam dumps →
JA
Written and reviewed by Johnson Ajibi, MSc IT Security
Senior Network & Security Engineer · founder of Courseiva
Last reviewed September 2026 · checked against the official Google Cloud exam blueprint
This PMLE practice question is part of Courseiva's free Google Cloud certification practice question bank. Courseiva provides original exam-style practice questions with explanations, topic-based practice, mock exams, readiness tracking, and study analytics to help learners prepare for the PMLE exam.