无法控制音量及每次克隆的音色不准确
Open
Nobody has claimed this yet.
- Dominant language
- Python
- Stars
- 37.8k
- Forks
- 4.3k
- Avg merge
- 7m
- Merged PRs (30d)
- 1
Description
- 没有办法通过一个字段控制音量,只能是提示词控制,不是很方便,如果第一次不理想,第二次也不理想,它没有记忆,音量稍微大一点儿和再大一点这个尺度不好拿捏,最好有个字段,是现在参考音频的音量的几倍,如3x,-1x等。
- 同一段文案,同一个参考音频,多次克隆有的时候太离谱,不是很理想,有的语速很快,有的时候音色根本和原音频一点不像,最好还有参数可以调整
Contributor guide
No contributing guide indexed for this repository
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Research direction
The issue does not name any files, entry points, or tests. Start by locating the voice-cloning generation and parameter-handling entry points, then inspect how repeated runs use the reference audio. Done should include controllable volume and cloning parameters, with tests or reproducible checks covering stable output behavior.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python
- Domain
- machine-learning
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Active
- Clarity
- Mostly clear
- Newbie friendliness
- 35/100