Courseiva

MLA-C01

Full exam simulation

2:10:00
1

ML Solution Monitoring, Maintenance, and Security

hard

A company deploys a real-time inference endpoint with auto-scaling using a target tracking policy based on average Invocations per instance. They notice that during a traffic spike, the endpoint scales out too late, causing increased latency. They want to scale proactively before the spike. Which strategy should they implement?

0 of 180 answered