mudler / mudler/LocalAI

TTS Voice Design WebUI

Open
#11,609 2 comments 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

enhancement
Dominant language
Go
Stars
49.2k
Forks
4.5k
Avg merge
1d 3m
Merged PRs (30d)
239

Description

Problem
Issue 10164 was resolved with a much needed feature for voice design, however the webui doesn’t appear to support this

Proposed Solution
Propose adding an option to the studio that allows sending an instructions parameter rather than as a fixed voice or adding an option to the operate->voice library for creating a voice using instructions rather than cloning only

Alternatives Considered
Using curl works but is not ideal for a fully built in solution

Contributor guide

Open the contributing guide

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

Start by reviewing the WebUI studio and the operate→voice flow described in the issue, then compare their requests with the existing voice-design API from issue 10164. Done means the WebUI can submit an instructions parameter for voice design without requiring curl, with the supported flow available from the built-in interface.

Written by the indexing model from the issue text.

Assessment

Domain
audio-video-rtc, frontend
Issue type
Feature
Difficulty
3/5
Estimated time
1-2 days
Activity status
Quiet
Clarity
Mostly clear
Newbie friendliness
52/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.