googleapis / googleapis/google-cloud-cpp

Adjust `end of speech` timeout on streaming speech to text

Ouverte
#15,225 0 commentaires 0 réactions 0 personnes assignées Voir sur GitHub
type: feature request
Langage dominant
C++
Étoiles
659
Forks
462
Merge moyen
1 j 2 h
PR mergées (30 j)
89

Description

**What component of `google-cloud-cpp` is this feature request for?**

google/cloud/speech/speech_client

**Is your feature request related to a problem? Please describe.**

I would like to reduce timeout it takes on Google before end of speech is detected

**Describe the solution you'd like**

I would like to have parameter added to streaming recognition config (speech::v1::StreamingRecognizeRequest) that would setup this timeout, similarly to what EndSilenceTimeout is on MSFT (https://learn.microsoft.com/en-us/windows/apps/design/input/set-speech-recognition-timeouts)

> End Silence Timeout:
> This timeout is triggered after a phrase has been successfully recognized, and the service waits for further speech input before finalizing the recognition result.
> The EndSilenceTimeout property determines how long the service waits for additional speech after a recognized phrase before concluding that the speech input has ended.
> This timeout can be adjusted to accommodate various speaking styles, allowing for short pauses within a longer phrase without prematurely ending the recognition.

**Describe alternatives you've considered**

No alternatives are known to exist, but please let me know otherwise

**Additional context**

Request originates from work on AI voice assistants where I would like to experiment with slightly reduced end of speech timeouts

Guide de contribution

Ouvrir le guide de contribution

Évaluation

Cette issue n'a pas encore été évaluée.

Recevez les nouvelles issues par e-mail

Un résumé court des issues GitHub adaptées aux débutants.