GoogleCloudPlatform / GoogleCloudPlatform/python-docs-samples

Speech API python client library method `streaming_recognize()` doesn't return a result with `is_final` set to True

未关闭
#12,570 1 条评论 0 个 reaction 已指派 1 人 已被 @engelke 认领 在 GitHub 查看
priority: p2 samples triage me type: bug
主要语言
Jupyter Notebook
星标
8.1k
派生
6.7k
平均合并
4 天 2 小时
30 天内合并 PR
16

描述

## In which file did you encounter the issue?
- https://github.com/GoogleCloudPlatform/python-docs-samples/blob/main/speech/microphone/transcribe_streaming_infinite.py

## Did you change the file? If so, how?
I have changed the line 299
```
config = speech.RecognitionConfig(
encoding=speech.RecognitionConfig.AudioEncoding.LINEAR16,
sample_rate_hertz=SAMPLE_RATE,
language_code="en-US",
max_alternatives=1,
)
```
into
```
config = speech.RecognitionConfig(
encoding=speech.RecognitionConfig.AudioEncoding.LINEAR16,
sample_rate_hertz=SAMPLE_RATE,
language_code="ja-JP",
max_alternatives=1,
)
```

## Describe the issue
streaming_recognize() method doesn't work properly. For even when a speaker utters a simple greeting like こんにちは in a very silent environment, this method returns StreamingRecognizeResponse object with a result of is_final == False. And always returns a result with is_final set to True exactly 90 seconds later.
this doesn't happen when the language is English or Chinese.
The same thing also happens in `transcribe_streaming_infinite_v2.py`

I am not sure if this report should be post here for this seems to be more internal problem.
Appreciate any help or clues anyways.

贡献指南

打开贡献指南

评估

这个 Issue 还没有评估数据。

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。