googleapis / googleapis/python-genai

Large PDF documents take too long to process

Đang mở
#809 5 bình luận 0 reaction 1 người được giao Được @Venkaiahbabuneelam nhận Xem trên GitHub
api: gemini-api priority: p2 type: bug
Ngôn ngữ chính
Python
Star
4k
Fork
1k
Merge trung bình
2 ngày 11 giờ
Pull request đã merge (30 ngày)
40

Mô tả

Although there's a limit of 1000 pages for a PDF document, it becomes very difficult to use it starting after 300 pages. It takes too long for the API to return an answer (>3 minutes). I tried caching it and that also didn't seem to decrease latency. I found that very weird, since including a large file at gemini.google.com and asking something is much faster, usually returning an answer in less than 30s. Why is it that a request containing a large PDF takes too long in the API? Any tips on how to optimize the speed?

#### Environment details

- Programming language: Python
- OS: MacOS
- Language runtime version: 3.11
- Package version: 1.12.1

#### Steps to reproduce

```
chunk_1 = gemini_client.files.get(name="large_doc_part_1.pdf")
chunk_2 = gemini_client.files.get(name="large_doc_part_2.pdf")
response = gemini_client.models.generate_content(
model="gemini-2.5-pro-preview-05-06",
contents=[chunk_1, chunk_2, "Tell me something interesting about this document"],
)
```

Hướng dẫn đóng góp

Mở hướng dẫn đóng góp

Đánh giá

Issue này chưa được đánh giá.

Nhận issue mới trong hộp thư của bạn

Bản tóm tắt ngắn những issue GitHub phù hợp với người mới.