aws / aws/amazon-sagemaker-examples
Multi threads with Sagemaker multi-model-endpoint
- Dominant language
- Jupyter Notebook
- Stars
- 11k
- Forks
- 7k
- Avg merge
- 8h 29m
- Merged PRs (30d)
- 8
Description
Hey,
I am using the sklearn multi-model-endpoint.
On the load test, I see that the endpoint serving one by one request and not more.
It is possible to manage more than **one** request (according to the instance type and the number of cores) at the moment per instance with Sagemaker? how?
Thanks
Contributor guide
Research direction
The issue names a scikit-learn SageMaker multi-model endpoint but no repository file or test. Start by locating the relevant endpoint example and reviewing how serving and request concurrency are configured. Done means establishing whether concurrent requests are supported for the stated instance and documenting the applicable configuration or limitation.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- aws
- Domain
- cloud, machine-learning
- Issue type
- Feature
- Difficulty
- 4/5
- Estimated time
- 3-5 days
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 25/100