PMLE Serving and Scaling Models • Set 5
PMLE Serving and Scaling Models Practice Test 5 — 15 questions with explanations. Free, no signup.
A healthcare company is using Vertex AI Endpoints to serve a medical imaging model that requires GPU acceleration. They need to optimize cost while ensuring that predictions are served with low latency during business hours (8 AM to 6 PM) and can tolerate higher latency during off-hours. Which configuration should they use?
Choose an answer to begin — your selection is scored in the full session.
15 questions · instant feedback and full explanations after every question.