bugbakery / bugbakery/transcribee

speaker identification from multitrack audio

Open
#607 0 comments 0 reactions 0 assignees View on GitHub
enhancement
Dominant language
TypeScript
Stars
515
Forks
39
Avg merge
19h 36m
Merged PRs (30d)
15

Description

In some scenarios the audio is available in a multitrack format (one speaker per microphone). We should have the ability to ingest this as both one audio file with multiple channels and as multiple audio channels with separate files and derive speaker identification from this.

Contributor guide

Open the contributing guide

Research direction

The issue names no files, tests, or entry points, so first trace the audio-ingestion and speaker-identification paths in the repository. Done should include support for one multichannel file and multiple per-speaker files, with speaker identification derived from both formats.

Written by the indexing model from the issue text.

Assessment

Tech stack
typescript
Domain
audio-video-rtc
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Quiet
Clarity
Needs clarification
Newbie friendliness
35/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.