功能建议:增加双向语音对话能力(语音输入 + 语音朗读)
- Dominant language
- TypeScript
- Stars
- 2.7k
- Forks
- 395
- Avg merge
- 21h 48m
- Merged PRs (30d)
- 776
Description
**提交人**: 普普通通Tony大叔
**客户端版本**: 0.1.58
---
## 需求描述
希望 Cindy 支持双向语音对话,既能接收用户的语音输入,也能把回复用语音朗读出来,覆盖手机端、桌面端与飞书等多种使用场景。
## 使用场景
- 随时随地使用,尤其是不便打字或希望解放双手的场景(如走路、通勤、做家务等)。
- 在手机端、桌面端、飞书内都希望获得一致的语音交互体验。
## 当前痛点
当前仅支持文本、图片、本地文件三种输入,没有语音通道。用户若想“说话”,需要先用系统语音输入法手动转成文字再发送,无法得到无缝的语音体验;回复也只能阅读文本。
## 诉求
- 语音输入:可将用户语音转换为文字供对话。
- 语音朗读:可将模型回复转换为语音播放。
- 两者结合实现真正的双向语音对话;并希望在各端(手机 / 桌面 / 飞书)保持一致。
## 建议方案
作为通用能力集成到主对话界面,并考虑覆盖常见客户端平台。
(以上为功能建议,非 bug;作为用户反馈提交。)
---
**版本区域**: CN
**OS**: win32 x64 (10.0.26200)
**界面语言**: zh-CN
Contributor guide
Research direction
The issue names no files, tests, or entry points. Begin by locating the main conversation UI and the existing integrations for mobile, desktop, and Feishu, then clarify the speech input and output boundaries and cross-platform completion criteria before implementation.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- typescript
- Domain
- desktop, frontend, mobile
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Active
- Clarity
- Needs clarification
- Newbie friendliness
- 25/100