MLA-C01 • Practice Test 8
Free MLA-C01 practice test — 10 questions with explanations. Set 8. No signup required.
A team is deploying a machine learning model using Amazon SageMaker. They need to serve predictions with sub-100ms latency for a real-time application. The model is a large ensemble that requires 4 GB of memory. The team expects traffic of 100 requests per second initially, but it may double during peak hours. Which instance type and deployment configuration should the team choose to minimize cost while meeting the latency requirement?
Choose an answer to begin — your selection is scored in the full session.
10 questions · instant feedback and full explanations after every question.