alphacep / alphacep/vosk-api

vosk can become very slow when spoken language is different from model

オープン
#660 コメント 9 件 リアクション 0 件 担当者 0 名 GitHub で見る
主要言語
Jupyter Notebook
スター
15.1k
フォーク
1.8k
PR マージ指標
30日以内にマージされた PR はありません

説明

When e.g. using vosk-model-en-us-daanzu-20200905 on a modern CPU (single thread), performance can be up to about 15x realtime. However, when the spoken language switches to something different from the model (and possibly with background noise) I see Vosk sometimes slowing down to .1x realtime or less. Is it possible to avoid this slow-down scenario somehow? If it is useful I can collect some audio samples that exhibit this but my guess is that this is fundamental to the implementation.

コントリビューションガイド

このリポジトリのコントリビューションガイドは索引されていません

評価

この issue はまだ評価されていません。

新しい issue をメールで受け取る

初心者向けの GitHub issue を短くまとめたダイジェスト。