microsoft / microsoft/ai-dev-gallery

[FEATURE] need text 2 voice model & voice 2 text

Open
#337 1 comment 2 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

sample
Dominant language
C#
Stars
1.5k
Forks
223
PR merge metrics
No merged PRs in 30d

Description

Is your feature request related to a problem? Please describe.
no

Describe the solution you'd like
need AI Speaker

Describe alternatives you've considered
support text 2 voice model(onnx) such as rhasspy/piper.
voice 2 text model need too.

Additional context

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

The issue names an AI Speaker, ONNX text-to-voice models such as rhasspy/piper, and a voice-to-text model, but it does not identify files, tests, entry points, or an acceptance criterion. Start by locating the project's existing model integrations and determine the intended speech-model scope before implementation; done would require an agreed design and working text-to-speech and speech-to-text support.

Written by the indexing model from the issue text.

Assessment

Domain
ai, audio-video-rtc, machine-learning
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
18/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.