alphacep / alphacep/vosk-api

using this stuff for a newbie

未关闭
#1,540 1 条评论 0 个 reaction 已指派 0 人 在 GitHub 查看
主要语言
Jupyter Notebook
星标
15.1k
派生
1.8k
PR 合并指标
30 天内没有已合并 PR

描述

hello, i'm new to speech recognition, vosx and python, but i want to translate speech from a simple video i downloaded from the internet (and later even tts'ing to my language or even speech to speech).
i have tried the listen_in_background function in the example with google engine and it works although i'm not able to obtain my goal (word by word translation)

with your software, the recognize_vosx in the callback keeps giving me the same result "Please download the model etc..." and i have done it and unzipped in vosk/model/it (i'm italian) but I can't get it to work.

so i have python3.12 installed, pyaudio, speech_recognition, and my ide for now is the simple IDLE, can you please give me a simple source to begin with this stuff?

贡献指南

这个仓库没有索引到贡献指南

评估

这个 Issue 还没有评估数据。

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。