RVC-Project / RVC-Project/Retrieval-based-Voice-Conversion-WebUI

Training on Apple Silicon

Open
#767 17 comments 3 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Dominant language
Python
Stars
38.4k
Forks
5.3k
PR merge metrics
No merged PRs in 30d

Description

Hey all,

I used my M1 MacBook Pro and trained a model with a dataset with .wav format that is 30 min long without slicing it, and I have some questions:

  1. Is it reasonable that it took more than 6 hours?
  2. Do slicing the dataset into < 10sec for every part, can make the training process faster?
  3. Is it that slow because RVC still doesn't support Apple Silicon, and the GPU is not being used?
  4. Are there any MPS updates coming to make training faster?

Hope some of you can help me with this 👯

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

The issue names no repository files, tests, or entry points. Start by reproducing the reported six-hour training run with the 30-minute WAV dataset and locating the training path and Apple Silicon/MPS handling; done requires a confirmed scope and answers or an implemented support plan.

Written by the indexing model from the issue text.

Assessment

Tech stack
macos, python
Domain
machine-learning, performance
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
15/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.