OpenBMB / OpenBMB/MiniCPM-o-Demo
MiniCPM-o 4.5 全双工模式,有较多时候说话结束,模型没有反馈,语音中反复强调说话结束,模型才有应答。不知是模型问题还是我的代码问题?
Open
Nobody has claimed this yet.
- Dominant language
- Python
- Stars
- 386
- Forks
- 81
- Avg merge
- 1h 59m
- Merged PRs (30d)
- 3
Description
模拟真人沟通,全双工问答场景。流式发送音频给模型,有时候能回馈应答,有时候音频没有人声了,模型依旧不应达,反复提示我依旧回答结束,模型才有文本和音频的反馈。不知是否是模型的能力问题?如果不是,可能能简单指出可能是什么原因?
Contributor guide
No contributing guide indexed for this repository
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Research direction
No file or test is identified. Start by reproducing the streaming audio interaction in the demo and trace the end-of-speech and response flow. Done means determining whether the missed response is reproducible in the model or caused by the demo's audio handling, with enough details for a targeted fix.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python
- Domain
- audio-video-rtc, machine-learning
- Issue type
- Bug
- Difficulty
- 4/5
- Estimated time
- 3-5 days
- Activity status
- Quiet
- Clarity
- Needs clarification
- Newbie friendliness
- 35/100