alphacep / alphacep/vosk-api

Utterance contains multiple speakers

Open
#488 2 comments 0 reactions 0 assignees View on GitHub
Dominant language
Jupyter Notebook
Stars
15.1k
Forks
1.8k
PR merge metrics
No merged PRs in 30d

Description

I'm noticing that I'm getting a lot of utterances that contain multiple speakers. This will give me speaker vectors that are calculated from multiple voices. This leads to a lot of bad matches. Any ideas on how I can get better results? Any way to tweak for better utterances between speakers? Any way for the speaker model to detect multiple speakers?

I'm using `python vosk-0.3.18` with `vosk-model-en-us-daanzu-20200905-lgraph` and `vosk-model-spk-0.4`

Thanks!

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.