googleapis / googleapis/google-cloud-cpp

Adjust `end of speech` timeout on streaming speech to text

Aperta
#15,225 0 commenti 0 reazioni 0 assegnatari Vedi su GitHub
type: feature request
Lingua principale
C++
Stelle
659
Fork
462
Merge medio
1g 2h
PR unite (30g)
89

Descrizione

**What component of `google-cloud-cpp` is this feature request for?**

google/cloud/speech/speech_client

**Is your feature request related to a problem? Please describe.**

I would like to reduce timeout it takes on Google before end of speech is detected

**Describe the solution you'd like**

I would like to have parameter added to streaming recognition config (speech::v1::StreamingRecognizeRequest) that would setup this timeout, similarly to what EndSilenceTimeout is on MSFT (https://learn.microsoft.com/en-us/windows/apps/design/input/set-speech-recognition-timeouts)

> End Silence Timeout:
> This timeout is triggered after a phrase has been successfully recognized, and the service waits for further speech input before finalizing the recognition result.
> The EndSilenceTimeout property determines how long the service waits for additional speech after a recognized phrase before concluding that the speech input has ended.
> This timeout can be adjusted to accommodate various speaking styles, allowing for short pauses within a longer phrase without prematurely ending the recognition.

**Describe alternatives you've considered**

No alternatives are known to exist, but please let me know otherwise

**Additional context**

Request originates from work on AI voice assistants where I would like to experiment with slightly reduced end of speech timeouts

Guida per i contributori

Apri la guida per i contributori

Valutazione

Questa issue non è ancora stata valutata.

Ricevi le nuove issue nella tua casella

Un breve riepilogo di issue GitHub adatte ai principianti.