aws / aws/amazon-sagemaker-examples

Multi threads with Sagemaker multi-model-endpoint

Open
#1,303 1 comment 0 reactions 0 assignees View on GitHub
Dominant language
Jupyter Notebook
Stars
11k
Forks
7k
Avg merge
8h 29m
Merged PRs (30d)
8

Description

Hey,
I am using the sklearn multi-model-endpoint.
On the load test, I see that the endpoint serving one by one request and not more.
It is possible to manage more than **one** request (according to the instance type and the number of cores) at the moment per instance with Sagemaker? how?

Thanks

Contributor guide

Open the contributing guide

Research direction

The issue names a scikit-learn SageMaker multi-model endpoint but no repository file or test. Start by locating the relevant endpoint example and reviewing how serving and request concurrency are configured. Done means establishing whether concurrent requests are supported for the stated instance and documenting the applicable configuration or limitation.

Written by the indexing model from the issue text.

Assessment

Tech stack
aws
Domain
cloud, machine-learning
Issue type
Feature
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.