aws / aws/amazon-sagemaker-feedback

SageMaker endpoint deployment fails despite available ml.m5.large quota

Open
#251 1 comment 0 reactions 0 assignees View on GitHub
bug inference-and-endpoints instances models reliability&stability servicequotas
Dominant language
No language data
Stars
10
Forks
3
PR merge metrics
No merged PRs in 30d

Description

### Product Version

- [ ] Amazon SageMaker Studio Classic
- [x] Amazon SageMaker Studio
- [ ] Issue is not related to SageMaker Studio

### Issue Description

I am trying to deploy a JumpStart model using an ml.m5.large endpoint.

Service Quotas shows an applied quota of 4 with utilization 0.

SageMaker also shows 0 of 4 quota used, but the deployment fails with the error: limit: 0, in use: 0, requested: 1.

Both SageMaker and Service Quotas are using the same AWS Region.

Steps to reproduce:

Open SageMaker JumpStart.

1. Select a model and choose Deploy.
2. Select ml.m5.large.
3. Set instance count to 1.
4. Click Deploy.
5. Deployment fails with quota limit 0, even though Service Quotas shows 4 available.

Image

### Expected Behavior

_No response_

### Observed Behavior

_No response_

### Product Category

JumpStart

### Feedback Category

Service Quotas

### Other Details

_No response_

Contributor guide

Open the contributing guide

Research direction

Start by reproducing the deployment in SageMaker Studio JumpStart with an ml.m5.large instance, then compare the quota shown in SageMaker with the same-region value in Service Quotas. Done means the deployment succeeds with the available quota or the service clearly reports and resolves the quota discrepancy.

Written by the indexing model from the issue text.

Assessment

Tech stack
aws
Domain
cloud, machine-learning
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Active
Clarity
Needs clarification
Newbie friendliness
42/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.