Persian model (vosk-model-small-fa-0.5) transcribes only 1–2 words from full speech
Ouverte
- Langage dominant
- Jupyter Notebook
- Étoiles
- 15.1k
- Forks
- 1.8k
- Métriques de merge des PR
- Aucune PR mergée en 30 j
Description
Hi,
I'm testing the Persian model vosk-model-small-fa-0.5 for an offline voice-to-text task. I noticed that the model performs very poorly on clear Persian audio input. Even when the input is a complete sentence spoken clearly in a quiet environment, the output typically includes only one or two unrelated words, or sometimes nothing at all.
Guide de contribution
Aucun guide de contribution indexé pour ce dépôt
Évaluation
Cette issue n'a pas encore été évaluée.