alphacep / alphacep/vosk-api

Persian model (vosk-model-small-fa-0.5) transcribes only 1–2 words from full speech

未關閉
#1,957 1 則留言 0 個 reaction 已指派 0 人 在 GitHub 檢視
主要語言
Jupyter Notebook
星號
15.1k
分支
1.8k
PR 合併指標
30 天內沒有已合併 PR

描述

Hi,
I'm testing the Persian model vosk-model-small-fa-0.5 for an offline voice-to-text task. I noticed that the model performs very poorly on clear Persian audio input. Even when the input is a complete sentence spoken clearly in a quiet environment, the output typically includes only one or two unrelated words, or sometimes nothing at all.

貢獻指南

這個儲存庫沒有索引到貢獻指南

評估

這個 Issue 還沒有評估資料。

把新 issue 寄到你的電子郵件信箱

精選適合新手參與的 GitHub issue 摘要。