MLS-C01 Modeling • Set 36
MLS-C01 Modeling Practice Test 36 — 15 questions with explanations. Free, no signup.
A company is deploying a real-time inference endpoint using SageMaker. The model is a large deep learning model (5 GB) with strict latency requirements (< 100 ms per request). The team expects bursty traffic with up to 1000 requests per second. Which configuration best meets the latency and throughput requirements?
Choose an answer to begin — your selection is scored in the full session.
15 questions · instant feedback and full explanations after every question.