alphacep / alphacep/vosk-api

Question about using multiple recognizers with same model and different grammar

Aperta
#1,102 2 commenti 0 reazioni 0 assegnatari Vedi su GitHub
Lingua principale
Jupyter Notebook
Stelle
15.1k
Fork
1.8k
Metriche di merge delle PR
Nessuna PR unita negli ultimi 30g

Descrizione

Hi Team,
I am using Vosk api for continous speech recognition in an android application. The offline ASR works very well and by supplying grammar word list the recognition has very high accuracy.
Currently, I want to let the app accept more complicated voice command from user. For example, when user first speak "open file", then I will ask him to provide the file name. In the second phase, I need to recognize the file name from his voice input. However, since all possible filenames is dynamically loaded from db, I will need to start a new recognizer (with the file names as grammar) and speechService.
So my question is whether dynamically construct multiple short-term used recognizer and speechservice be expensive on resource. Or will that cause problem if I have several recognizer/services created in memory and switch between them based on application context?
Thank you for any suggestion.
Regards,
Steven

Guida per i contributori

Nessuna guida per i contributori indicizzata per questo repository

Valutazione

Questa issue non è ancora stata valutata.

Ricevi le nuove issue nella tua casella

Un breve riepilogo di issue GitHub adatte ai principianti.