A model deployed to a SageMaker serverless endpoint shows higher latency in production than in tests. The engineer suspects the extra latency is caused by model startup time. What should they check to verify this?
Choose an answer
Tap an option to check your answer.
Correct answer: Enable Amazon CloudWatch metrics and inspect the ModelSetupTime metric in the SageMaker namespace..
Why this is the answer
The correct answer is to enable Amazon CloudWatch metrics and inspect the ModelSetupTime metric in the SageMaker namespace. This metric specifically measures the time taken for the model container to initialize and load the model artifacts, which directly correlates with model startup time. If this value is high, it confirms the engineer's suspicion. Scheduling a SageMaker Model Monitor job (with or without CloudWatch metrics) is incorrect because Model Monitor focuses on data quality, model quality, and bias drift, not on the operational performance metric of model startup time. Inspecting the ModelLoadingWaitTime metric is incorrect. While related to loading, ModelSetupTime is the more encompassing and direct metric for the overall startup duration of the model within the container for a serverless endpoint.
Pass your exam — without the endless answer hunt
Get every verified question and explanation for this exam in one place, and save hours of prep. 1,000+ certifications · 20+ languages · free to start.
Pass your exam faster → No card needed