alphacep / alphacep/vosk-api

Training a model that will be used only in grammar mode

Ouverte
#1,635 1 commentaire 0 réactions 0 personnes assignées Voir sur GitHub
Langage dominant
Jupyter Notebook
Étoiles
15.1k
Forks
1.8k
Métriques de merge des PR
Aucune PR mergée en 30 j

Description

When training a model for other language that will be used only in grammar mode, do I need less data compared to models being trained for dictation? If so, can you estimate how many hours of audio do I need?

In my use case the words that will be used in the grammar and not known to me and will be entered by the user of the app. Also, the grammar should include "unk" for unrecognized words.

Thanks

Guide de contribution

Aucun guide de contribution indexé pour ce dépôt

Évaluation

Cette issue n'a pas encore été évaluée.

Recevez les nouvelles issues par e-mail

Un résumé court des issues GitHub adaptées aux débutants.