AWS Certified Machine Learning Engineer - Associate MLA-C01 — Question 158
Topic 1 · Question 158 of 226
Topic 1 · Question 158
A company uses an Amazon SageMaker AI ML model to make real-time inferences. The company has configured auto scaling for the Amazon EC2 instances that SageMaker AI uses for the inferences. During times of peak usage, new instances launch before existing instances are fully ready. As a result, the model experiences inefficiencies and delays. Which solution will optimize the scaling process without affecting response times?