OpenBMB / OpenBMB/VoxCPM

英文朗读的问题

Open
#126 2 comments 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Dominant language
Python
Stars
37.8k
Forks
4.3k
Avg merge
7m
Merged PRs (30d)
1

Description

问题:
  1. 为什么中英文混读的时候,英文的劣化这么严重?经常连基本的单词都读不明白。这是这版本目前测试下来最大的问题。

纯英文场景我没测试过,但是在中文中穿插着单词,哪怕这是个常见的标准单词,依然有很高的概率出现读法完全错误。

这是提示语音的问题还是模型的问题?

说明:
我的提示语音使用sensevoice small识别出来的文本,语音朗读是很标准的普通话。

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

No source file or test is identified. Start by reproducing mixed Chinese-English speech with the SenseVoice Small transcription and standard Mandarin prompt described in the report, then compare the prompt text and model output; done means determining whether the degradation comes from the prompt or the speech model.

Written by the indexing model from the issue text.

Assessment

Tech stack
python, pytorch
Domain
machine-learning
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
30/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.