aws / aws/amazon-sagemaker-examples

[Example Request] Is it possible to use SageMaker for streaming speech recognition?

Open
#2,940 1 comment 0 reactions 0 assignees View on GitHub
Dominant language
Jupyter Notebook
Stars
11k
Forks
7k
Avg merge
8h 29m
Merged PRs (30d)
8

Description

Hi @javier @tomfaulhaber @channy @danny @ishaaq ,
We are working on a custom speech-to-text project, and we need to deploy the real-time speech-to-text server in the cloud. While checking the SageMaker, we found few limitations like 60 seconds timeout. So we have some doubts regarding this.
For real-time speech to text (audio-stream to text), can we use SageMaker?
Using Lambda + ApiGateway (WebSocket) for audio-stream and pass the audio to SageMaker. Is it a good solution?

Other than SageMaker, we are considering using the ECS cluster with auto-scaling, Will it take time to initialize the new cluster while auto-scaling?.

If anybody hosted a real-time speech recognition server on AWS, please help us decide on a solution.

Let us know if there is any better stack for real-time speech recognition stack we can use in AWS.

Contributor guide

Open the contributing guide

Research direction

The issue discusses SageMaker, Lambda, API Gateway WebSocket, and ECS for real-time speech recognition, but names no repository files, tests, or entry points. First clarify whether an AWS example is requested and which architecture should be demonstrated; the issue provides no concrete completion criterion.

Written by the indexing model from the issue text.

Assessment

Tech stack
aws
Domain
cloud, machine-learning
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
15/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.