Quiz collegato alla guida · Medio Livello

Autoscaling Model Inference on Kubernetes Quiz

Configure model-serving autoscaling signals and account for GPU scheduling, model startup delays, queues and scale-to-zero tradeoffs.

Domanda 1 di 8

Which signal may represent inference pressure better than CPU alone?