OpenMOSS / OpenMOSS/MOSS-TTSD

对话语音有杂音

Open
#83 0 comments 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Dominant language
Python
Stars
1.4k
Forks
138
PR merge metrics
No merged PRs in 30d

Description

在GITEE的模型广场, https://ai.gitee.com/serverless-api 选择 MOSS-TTSD-v0.5 在线体验,默认不改动,直接点击运行,生成的对话有杂音。自己用API换文字,没换参考音频,也是有对话杂音,是需要哪里设置一下么?

FLSL1RU4N5QMMGAZQDBRERUSUTHBS91U.wav

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

No source files or tests are named in the issue. Start by reproducing the default MOSS-TTSD-v0.5 run at the linked GITEE model page and then the API case using the attached WAV as the symptom reference. Done means the source of the noise is identified and a clear configuration or code change is specified.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
audio-video-rtc, machine-learning
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
35/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.