huggingface / huggingface/diarizers

How do I retrive the speech text?

Open
#5 1 comment 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
331
Forks
22
PR merge metrics
No merged PRs in 30d

Description

Running starter program, it detects the speakers. How do I get the speech?

My usecase is I like to output a transcript in following form

Start_time, End_Time, SpeechText, Speaker_id

Contributor guide

No contributing guide indexed for this repository

Research direction

Start with the starter program referenced in the issue and inspect how it exposes speaker-detection results and whether speech text is available. Done means the repository documents or demonstrates obtaining Start_time, End_time, SpeechText, and Speaker_id, or clearly explains if speech transcription is outside the program's scope.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
audio-video-rtc
Issue type
Documentation
Difficulty
3/5
Estimated time
1-2 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.