babysor / babysor/MockingBird

生成的音频开头和结尾的停顿和渐进如何去除?

Open
#896 1 comment 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
36.9k
Forks
5.2k
PR merge metrics
No merged PRs in 30d

Description

您好,我是小白,有两个问题想要请教一下。
① 目前只能生成一些短的句子,超过十个字后面的就基本都是噪音,请问这个是有什么方法可以改进的吗?
② 生成的音频开头和结尾都有短暂的停顿,导致拼接起来的时候中间比较不连续,而且开头和结尾的音量是渐进的,开头声音逐渐变大,结尾逐渐变小,请问我如果不想要这个效果可以怎么改呢?
谢谢!

Contributor guide

No contributing guide indexed for this repository

Research direction

No file, test, or entry point is named. Start by reproducing both reported behaviors with text longer than ten Chinese characters and inspect the generated audio for noise, leading or trailing pauses, and volume fades. Done would require a confirmed cause and documented change or configuration that addresses the reported generation behavior.

Written by the indexing model from the issue text.

Assessment

Tech stack
python, pytorch
Domain
ai, machine-learning
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.