babysor / babysor/MockingBird

跑的好像收敛了,试听也还行,但是打开程序合成音频效果很差,是不是因为没有解压aidatatang_200zh的数据

Open
#776 2 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
36.9k
Forks
5.2k
PR merge metrics
No merged PRs in 30d

Description

要不要解压aidatatang_200zh数据之后重新预处理然后训练,我自己的数据有350条,识别出了261条,时长应该有10分钟
前面有一两次损失是0.4附近的

![attention_step_78500_sample_1](https://user-images.githubusercontent.com/117540847/200137993-44261951-7cbb-4621-888c-36cc07a33baa.png)

![step-78500-mel-spectrogram_sample_1](https://user-images.githubusercontent.com/117540847/200137999-bc8fe0e4-562e-4f62-8410-0c79eccbc469.png)

Contributor guide

No contributing guide indexed for this repository

Research direction

No file, test, or entry point is identified. First verify whether aidatatang_200zh is extracted before preprocessing, then compare the resulting training data and audio quality; done means determining whether extraction and reprocessing account for the poor synthesized audio.

Written by the indexing model from the issue text.

Assessment

Tech stack
python, pytorch
Domain
ai, machine-learning
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.