PMLE Monitoring ML Solutions • Set 2
PMLE Monitoring ML Solutions Practice Test 2 — 15 questions with explanations. Free, no signup.
A financial services company deploys a model on Vertex AI Endpoints with GPU acceleration. They notice that the p99 latency for predictions has increased from 200ms to 1.2s over the past week. CPU utilisation is low, but GPU utilisation is high. Which action should they take to reduce latency?
Choose an answer to begin — your selection is scored in the full session.
15 questions · instant feedback and full explanations after every question.