alphacep / alphacep/vosk-api

German small model and the umlaut

未关闭
#2,040 1 条评论 0 个 reaction 已指派 0 人 在 GitHub 查看
主要语言
Jupyter Notebook
星标
15.1k
派生
1.8k
PR 合并指标
30 天内没有已合并 PR

描述

Hello, folks.

I am trying out the small German model "vosk-model-small-de-0.15" by passing in grammar to the recognizer. What I am experiencing is that it works great, except for when a word that is passed in grammar contains an umlaut.

For instance, passing in the following json:
"[\"eins\",\"zwei\",\"drei\",\"vier\",\"fünf\",\"sechs\",\"sieben\",\"acht\",\"neun\",\"zehn\",\"elf\",\"zwölf\"]"

The recognizer will respond perfectly with everything but "fünf" and "zwölf".

I tried with the large model "vosk-model-de-0.21", and it seems to work with words that contain an umlaut. Is there something I am missing in regard to the small model or is this a known limitation? I haven't been able to find information on this, so I am hoping it is just me ;)

Thank you for all of your hard work on this.

贡献指南

这个仓库没有索引到贡献指南

评估

这个 Issue 还没有评估数据。

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。