RVC-Project / RVC-Project/Retrieval-based-Voice-Conversion-WebUI

Why does my RVC model produces bad results?

Open
#2,025 1 comment 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

help wanted
Dominant language
Python
Stars
38.4k
Forks
5.3k
PR merge metrics
No merged PRs in 30d

Description

I created a RVC model and I had it set to 50 epochs for training. It produces bad results. Here is the audio file I used for training, I used version v2 and I set the target sample rate to 48k.
https://files.catbox.moe/440g5r.wav
Here is the input file I used.
https://vocaroo.com/11ANq5W9xFd5
Here is the output file.
https://vocaroo.com/136PewYoZmL3

Edit: I discovered Seed-VC which worked well with 6 minutes of training data. It was also worked with 14 seconds of training data.

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

Start by reviewing the linked training, input, and output audio along with the v2, 48kHz, and 50-epoch settings. No source file or test is identified, so first establish a reproducible failure and define the expected output quality; the issue is complete only when the cause and a concrete corrective change are clear.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
audio-video-rtc, machine-learning
Issue type
Bug
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
15/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.