alphacep / alphacep/vosk-api

Training a model that will be used only in grammar mode

Offen
#1,635 1 Kommentar 0 Reaktionen 0 zugewiesene Personen Auf GitHub ansehen
Vorherrschende Sprache
Jupyter Notebook
Sterne
15.1k
Forks
1.8k
PR-Merge-Kennzahlen
Keine gemergten PRs in 30 T.

Beschreibung

When training a model for other language that will be used only in grammar mode, do I need less data compared to models being trained for dictation? If so, can you estimate how many hours of audio do I need?

In my use case the words that will be used in the grammar and not known to me and will be entered by the user of the app. Also, the grammar should include "unk" for unrecognized words.

Thanks

Beitragsleitfaden

Für dieses Repository ist kein Beitragsleitfaden indexiert

Bewertung

Dieses Issue wurde noch nicht bewertet.

Neue Issues direkt in Ihr Postfach

Eine kurze Übersicht über anfängerfreundliche GitHub-Issues.