MLA-C01 Deployment and Orchestration of ML Workflows • Set 2
MLA-C01 Deployment and Orchestration of ML Workflows Practice Test 2 — 15 questions with explanations. Free, no signup.
A data science team needs to deploy a PyTorch model that performs real-time inference with sub-100ms latency. The model requires GPU acceleration, but the team wants to minimize cost by sharing GPU instances across multiple models. Which SageMaker hosting option should they choose?
Choose an answer to begin — your selection is scored in the full session.
15 questions · instant feedback and full explanations after every question.