You recently deployed a scikit-learn model to a Vertex AI endpoint. You are now testing the model on live production traffic. While monitoring the endpoint, you discover twice as many requests per hour than expected throughout the day. You want the endpoint to efficiently scale when the demand increases in the future to prevent users from experiencing high latency. What should you do?
fitri001
Highly Voted 9 months, 3 weeks agofitri001
9 months, 3 weeks agoguilhermebutzke
Most Recent 1 year agoYan_X
1 year agob1a8fae
1 year ago36bdc1e
1 year, 1 month agopikachu007
1 year, 1 month agosonicclasps
1 year ago