babysor / babysor/MockingBird

用预训练的模型合成的语音问题

Open
#164 1 comment 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
36.9k
Forks
5.2k
PR merge metrics
No merged PRs in 30d

Description

用的@miven的预训练模型,启动web,上传wav音频文件,合成的语音全是杂音,怎么解决

Contributor guide

No contributing guide indexed for this repository

Research direction

Start at the web entry point that accepts the uploaded WAV and invokes the pretrained @miven model. Reproduce the noisy synthesis with the stated model and audio input, then trace the model and preprocessing path; done means identifying the cause and confirming that the same workflow produces intelligible speech.

Written by the indexing model from the issue text.

Assessment

Tech stack
python, pytorch
Domain
machine-learning
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.