GPU search time for SageMaker Async Inference and general GPU availability.

0

Background I want to build an ML Inference pipeline that will use SageMaker Asynchronous Inference. To decrease costs I want to down all SageMaker Async Inference-related EC2s when no jobs are waiting (for example for time out of business hours or during working hours where there are no requests from my users).

The questions

  1. On average, how long does it take for AWS SageMaker Async Inference to get an up-and-running EC2 with a GPU ready to execute my ML tasks/inference?
  2. What is the current availability of GPU machines on AWS? Is there any shortage?
asked a month ago73 views
No Answers

You are not logged in. Log in to post an answer.

A good answer clearly answers the question and provides constructive feedback and encourages professional growth in the question asker.

Guidelines for Answering Questions