A company wants to monitor the cost of their Vertex AI prediction endpoint. They are charged per hour per replica and per request for GPU instances. Which approach should they use to track these costs?
Budget alerts and exporting billing data are standard practices for cost monitoring.
Why this answer
Vertex AI prediction costs are tracked via Cloud Billing. Budget alerts can be set up to notify when spending exceeds a threshold. Cost breakdown can be viewed in the Billing reports.