alphacep / alphacep/vosk-api

German small model and the umlaut

Ouverte
#2,040 1 commentaire 0 réactions 0 personnes assignées Voir sur GitHub
Langage dominant
Jupyter Notebook
Étoiles
15.1k
Forks
1.8k
Métriques de merge des PR
Aucune PR mergée en 30 j

Description

Hello, folks.

I am trying out the small German model "vosk-model-small-de-0.15" by passing in grammar to the recognizer. What I am experiencing is that it works great, except for when a word that is passed in grammar contains an umlaut.

For instance, passing in the following json:
"[\"eins\",\"zwei\",\"drei\",\"vier\",\"fünf\",\"sechs\",\"sieben\",\"acht\",\"neun\",\"zehn\",\"elf\",\"zwölf\"]"

The recognizer will respond perfectly with everything but "fünf" and "zwölf".

I tried with the large model "vosk-model-de-0.21", and it seems to work with words that contain an umlaut. Is there something I am missing in regard to the small model or is this a known limitation? I haven't been able to find information on this, so I am hoping it is just me ;)

Thank you for all of your hard work on this.

Guide de contribution

Aucun guide de contribution indexé pour ce dépôt

Évaluation

Cette issue n'a pas encore été évaluée.

Recevez les nouvelles issues par e-mail

Un résumé court des issues GitHub adaptées aux débutants.