PMLE Serving and Scaling Models • Set 8
PMLE Serving and Scaling Models Practice Test 8 — 15 questions with explanations. Free, no signup.
Your team has deployed a model on Vertex AI endpoints. You need to monitor the prediction latency to ensure it meets a 99th percentile SLO of 500ms. You want to set up an alert if the latency exceeds this threshold. Which metric should you use?
Choose an answer to begin — your selection is scored in the full session.
15 questions · instant feedback and full explanations after every question.