alphacep / alphacep/vosk-api

Persian model (vosk-model-small-fa-0.5) transcribes only 1–2 words from full speech

Aperta
#1,957 1 commento 0 reazioni 0 assegnatari Vedi su GitHub
Lingua principale
Jupyter Notebook
Stelle
15.1k
Fork
1.8k
Metriche di merge delle PR
Nessuna PR unita negli ultimi 30g

Descrizione

Hi,
I'm testing the Persian model vosk-model-small-fa-0.5 for an offline voice-to-text task. I noticed that the model performs very poorly on clear Persian audio input. Even when the input is a complete sentence spoken clearly in a quiet environment, the output typically includes only one or two unrelated words, or sometimes nothing at all.

Guida per i contributori

Nessuna guida per i contributori indicizzata per questo repository

Valutazione

Questa issue non è ancora stata valutata.

Ricevi le nuove issue nella tua casella

Un breve riepilogo di issue GitHub adatte ai principianti.