alphacep / alphacep/vosk-tts

Specifying reference text in GPT-SoVITS leads to messy audio output

Aberta
#43 2 comentários 0 reações 0 responsáveis Ver no GitHub
Linguagem predominante
Python
Estrelas
274
Forks
35
Métricas de merge de PRs
Nenhum PR com merge em 30d

Descrição

Hi there!

When playing around with your GPT-SoVITS fork I found that specifying reference text as the second argument to `get_tts_wav` function here:

https://github.com/alphacep/vosk-tts/blob/89b23a8b033133e25e3e7f53d07939645b8ea51c/training/gpt-sovits/inference_cli.py#L273

results in the output audio being either gibberish or completely silent.

Without specifying the reference text inference works as intended.

Is this an expected behavoiur?

Guia de contribuição

Nenhum guia de contribuição indexado para este repositório

Avaliação

Esta issue ainda não foi avaliada.

Receba novas issues na sua caixa de entrada

Um resumo curto de issues do GitHub para quem está começando.