OpenMOSS / OpenMOSS/MOSS-TTSD

在测试混合朗读功能时,英文的“中文口音”过重,听着有点滑稽

Open
#36 0 comments 1 reaction 0 assignees View on GitHub

Nobody has claimed this yet.

Dominant language
Python
Stars
1.4k
Forks
138
PR merge metrics
No merged PRs in 30d

Description

使用的gradio demo的第一个例子,然后修改例子的文本为:[S2]我们可以通过 description、button_style 和 icon 属性来增强按钮的视觉效果,也可以通过如 on_click() 这样的方法来实现其功能。暂时我们将功能部分保存到下面的相关章节中

Image

生成的音频:
tmpinjgti3k.zip

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

Start by reproducing the first example in the Gradio demo with the provided mixed Chinese-English text, then compare the generated audio with the attached sample. Trace the text-to-speech path used by that example and determine what would count as improved English pronunciation without degrading the Chinese speech. Done means the mixed-language output no longer has the reported heavy Chinese accent.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
machine-learning
Issue type
Bug
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.