Documentation of speaker identification/diarization capabilities
- Vorherrschende Sprache
- Jupyter Notebook
- Sterne
- 15.1k
- Forks
- 1.8k
- PR-Merge-Kennzahlen
- Keine gemergten PRs in 30 T.
Beschreibung
Is there any additional documentation or description of the Python test files? Some are pretty obvious. Some are not.
Transcript_scp.py is in a separate directory, python/test. What does that one do? There doesn't seem to be a sample input files it can use
And then there is test_srt and test_speaker. Not sure what test_srt is doing. Test_speaker looks like it’s a way to identify who is talking. Is there any additional documentation on that? Is it simply estimating relative distances of the speaker? There’s a big hard-code array in there. Not sure if that's something I would need
Beitragsleitfaden
Für dieses Repository ist kein Beitragsleitfaden indexiert
Bewertung
Dieses Issue wurde noch nicht bewertet.