huggingface / huggingface/diarizers
How do I retrive the speech text?
Open
- Dominant language
- Python
- Stars
- 331
- Forks
- 22
- PR merge metrics
- No merged PRs in 30d
Description
Running starter program, it detects the speakers. How do I get the speech?
My usecase is I like to output a transcript in following form
Start_time, End_Time, SpeechText, Speaker_id
Contributor guide
No contributing guide indexed for this repository
Research direction
Start with the starter program referenced in the issue and inspect how it exposes speaker-detection results and whether speech text is available. Done means the repository documents or demonstrates obtaining Start_time, End_time, SpeechText, and Speaker_id, or clearly explains if speech transcription is outside the program's scope.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python
- Domain
- audio-video-rtc
- Issue type
- Documentation
- Difficulty
- 3/5
- Estimated time
- 1-2 days
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 25/100