googleapis / googleapis/python-genai

Large PDF documents take too long to process

未关闭
#809 5 条评论 0 个 reaction 已指派 1 人 已被 @Venkaiahbabuneelam 认领 在 GitHub 查看
api: gemini-api priority: p2 type: bug
主要语言
Python
星标
4k
派生
1k
平均合并
2 天 12 小时
30 天内合并 PR
41

描述

Although there's a limit of 1000 pages for a PDF document, it becomes very difficult to use it starting after 300 pages. It takes too long for the API to return an answer (>3 minutes). I tried caching it and that also didn't seem to decrease latency. I found that very weird, since including a large file at gemini.google.com and asking something is much faster, usually returning an answer in less than 30s. Why is it that a request containing a large PDF takes too long in the API? Any tips on how to optimize the speed?

#### Environment details

- Programming language: Python
- OS: MacOS
- Language runtime version: 3.11
- Package version: 1.12.1

#### Steps to reproduce

```
chunk_1 = gemini_client.files.get(name="large_doc_part_1.pdf")
chunk_2 = gemini_client.files.get(name="large_doc_part_2.pdf")
response = gemini_client.models.generate_content(
model="gemini-2.5-pro-preview-05-06",
contents=[chunk_1, chunk_2, "Tell me something interesting about this document"],
)
```

贡献指南

打开贡献指南

评估

这个 Issue 还没有评估数据。

把新 issue 发到你的邮箱

精选适合新手参与的 GitHub issue 摘要。