alphacep / alphacep/vosk-api

Documentation of speaker identification/diarization capabilities

Aperta
#405 8 commenti 0 reazioni 0 assegnatari Vedi su GitHub
help wanted
Lingua principale
Jupyter Notebook
Stelle
15.1k
Fork
1.8k
Metriche di merge delle PR
Nessuna PR unita negli ultimi 30g

Descrizione

Is there any additional documentation or description of the Python test files? Some are pretty obvious. Some are not.

Transcript_scp.py is in a separate directory, python/test. What does that one do? There doesn't seem to be a sample input files it can use

And then there is test_srt and test_speaker. Not sure what test_srt is doing. Test_speaker looks like it’s a way to identify who is talking. Is there any additional documentation on that? Is it simply estimating relative distances of the speaker? There’s a big hard-code array in there. Not sure if that's something I would need

Guida per i contributori

Nessuna guida per i contributori indicizzata per questo repository

Valutazione

Questa issue non è ancora stata valutata.

Ricevi le nuove issue nella tua casella

Un breve riepilogo di issue GitHub adatte ai principianti.