Courseiva

MLA-C01

Full exam simulation

2:10:00
1

Deployment and Orchestration of ML Workflows

medium

A machine learning team has a model that needs to serve predictions with very low latency (under 10 ms) for a real-time web application. The model is a small ensemble of three neural networks that fits in memory. Which SageMaker inference option is MOST appropriate?

0 of 50 answered