A company is deploying a microservices application on Google Kubernetes Engine (GKE). The architect needs to ensure that the cluster can automatically scale nodes based on pod resource requests and that pods are scheduled efficiently across nodes. The company also wants to minimize costs by scaling down when demand is low. Which two configurations should the architect implement? (Choose two.)
Cluster Autoscaler automatically adjusts the number of nodes in a node pool based on the resource requests of pending pods. It scales up when pods cannot be scheduled due to insufficient resources and scales down when nodes are underutilized. Setting a minimum and maximum node count ensures cost control and availability. This directly addresses the need to scale nodes based on pod demands and minimize costs during low demand.
Why this answer
Cluster Autoscaler scales the number of nodes in a node pool based on pending pod resource requests, and setting pod resource requests ensures that the scheduler and autoscaler have accurate information to make scaling decisions. Together, they enable automatic node scaling and efficient scheduling while allowing scale-down to reduce costs. The other options either address pod scaling, add unnecessary complexity, or improve availability without meeting the core requirements.
Exam trap
The trap here is assuming that Horizontal Pod Autoscaler alone can scale nodes; it only scales pod replicas, not the underlying node pool.