bugbakery / bugbakery/transcribee
speaker identification from multitrack audio
Open
enhancement
- Dominant language
- TypeScript
- Stars
- 515
- Forks
- 39
- Avg merge
- 19h 36m
- Merged PRs (30d)
- 15
Description
In some scenarios the audio is available in a multitrack format (one speaker per microphone). We should have the ability to ingest this as both one audio file with multiple channels and as multiple audio channels with separate files and derive speaker identification from this.
Contributor guide
Research direction
The issue names no files, tests, or entry points, so first trace the audio-ingestion and speaker-identification paths in the repository. Done should include support for one multichannel file and multiple per-speaker files, with speaker identification derived from both formats.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- typescript
- Domain
- audio-video-rtc
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Quiet
- Clarity
- Needs clarification
- Newbie friendliness
- 35/100