OpenBMB / OpenBMB/MiniCPM-o-Demo

MiniCPM-o 4.5 全双工模式,有较多时候说话结束,模型没有反馈,语音中反复强调说话结束,模型才有应答。不知是模型问题还是我的代码问题?

Open
#28 1 comment 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Dominant language
Python
Stars
386
Forks
81
Avg merge
1h 59m
Merged PRs (30d)
3

Description

模拟真人沟通,全双工问答场景。流式发送音频给模型,有时候能回馈应答,有时候音频没有人声了,模型依旧不应达,反复提示我依旧回答结束,模型才有文本和音频的反馈。不知是否是模型的能力问题?如果不是,可能能简单指出可能是什么原因?

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

No file or test is identified. Start by reproducing the streaming audio interaction in the demo and trace the end-of-speech and response flow. Done means determining whether the missed response is reproducible in the model or caused by the demo's audio handling, with enough details for a targeted fix.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
audio-video-rtc, machine-learning
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Quiet
Clarity
Needs clarification
Newbie friendliness
35/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.