babysor / babysor/MockingBird

只有 CPU的情況, 感覺 web 效果比 toolbox 好一點

Open
#362 4 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
36.9k
Forks
5.2k
PR merge metrics
No merged PRs in 30d

Description

web 和 toolbox 都可以運行
但感覺 web 效果比 toolbox 好一點

toolbox 的電流音比較明顯而多
而且有時候輸入的是男聲, 他輸出的變成女聲了

已有下載 dataset, 是用最新版本, 用 pretrained的 data.

相似度仍是有很大的距離,是要自己 train 嗎?

輸入的 wav 再長一點有用嗎? 通常多久最好?

Contributor guide

No contributing guide indexed for this repository

Research direction

Start with the web and toolbox CPU execution paths and reproduce the reported electrical noise and voice-gender mismatch using the pretrained dataset and WAV input described. Compare their outputs and document whether the difference is reproducible, including input duration and whether training is required; the issue is done when the cause or a concrete reproduction and next step is established.

Written by the indexing model from the issue text.

Assessment

Tech stack
python, pytorch
Domain
machine-learning
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.