aws / aws/amazon-sagemaker-examples
"Host a Pretrained Model on SageMaker" tutorial fails with "Received server error (500) from primary and could not load the entire response body."
- Dominant language
- Jupyter Notebook
- Stars
- 11k
- Forks
- 7k
- Avg merge
- 8h 29m
- Merged PRs (30d)
- 8
Description
**Link to the notebook**
https://github.com/aws/sagemaker-python-sdk/files/14076220/deploy-v4-any-pytorch-tutorial.md
**Describe the bug**
"Host a Pretrained Model on SageMaker" tutorial fails with "An error occurred (ModelError) when calling the InvokeEndpoint operation: Received server error (500) from primary and could not load the entire response body."
**To reproduce**
follow https://sagemaker-examples.readthedocs.io/en/latest/sagemaker-script-mode/pytorch_bert/deploy_bert_outputs.html#Deploy-Model
**Logs**
# Expected behavior

# error seeing

# similar post
I made a similar post under aws/sagemaker-python-sdk: https://github.com/aws/sagemaker-python-sdk/issues/4395
Contributor guide
Research direction
Start with the linked deploy-v4-any-pytorch-tutorial.md notebook and follow the Deploy Model section at the referenced pytorch_bert tutorial. Compare the expected behavior with the ModelError response and review the related sagemaker-python-sdk issue for context. Done means the tutorial's endpoint invocation completes successfully and matches the expected result.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- aws, jupyter-notebook, pytorch
- Domain
- cloud, machine-learning
- Issue type
- Bug
- Difficulty
- 4/5
- Estimated time
- 3-5 days
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 30/100