RVC-Project / RVC-Project/Retrieval-based-Voice-Conversion-WebUI

Can you support an open-source project pyanno.audio that can separate the voices of multiple speakers?

Open
#282 0 comments 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

enhancement help wanted
Dominant language
Python
Stars
38.4k
Forks
5.3k
PR merge metrics
No merged PRs in 30d

Description

Hi,
I was wondering if it would be possible to support the voice separation function of pyanno.audio in future versions? This package would make file organization much easier, especially when there are multiple voices speaking in a single audio clip. Using this package would allow us to separate the voices into multiple files.
Here's the link to pyanno.audio's Github: https://github.com/pyannote/pyannote-audio
I have researched it extensively but found it difficult due to my lack of knowledge in Python. I have also been unable to find GUI versions of this package.
Thank you.

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

No repository files, tests, or entry points are identified in the issue. Start by locating the audio processing workflow and reviewing pyannote.audio's speaker-separation requirements; the work is complete only when multiple voices can be separated into distinct files through the project's supported interface.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
audio-video-rtc, machine-learning
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.