alphacep / alphacep/vosk-api

Utterance contains multiple speakers

オープン
#488 コメント 2 件 リアクション 0 件 担当者 0 名 GitHub で見る
主要言語
Jupyter Notebook
スター
15.1k
フォーク
1.8k
PR マージ指標
30日以内にマージされた PR はありません

説明

I'm noticing that I'm getting a lot of utterances that contain multiple speakers. This will give me speaker vectors that are calculated from multiple voices. This leads to a lot of bad matches. Any ideas on how I can get better results? Any way to tweak for better utterances between speakers? Any way for the speaker model to detect multiple speakers?

I'm using `python vosk-0.3.18` with `vosk-model-en-us-daanzu-20200905-lgraph` and `vosk-model-spk-0.4`

Thanks!

コントリビューションガイド

このリポジトリのコントリビューションガイドは索引されていません

評価

この issue はまだ評価されていません。

新しい issue をメールで受け取る

初心者向けの GitHub issue を短くまとめたダイジェスト。