googleapis / googleapis/python-genai

Large PDF documents take too long to process

Abierto
#809 5 comentarios 0 reacciones 1 asignado Reclamado por @Venkaiahbabuneelam Ver en GitHub
api: gemini-api priority: p2 type: bug
Lenguaje dominante
Python
Estrellas
4k
Forks
1k
Merge medio
2 d 11 h
PR fusionados (30 d)
40

Descripción

Although there's a limit of 1000 pages for a PDF document, it becomes very difficult to use it starting after 300 pages. It takes too long for the API to return an answer (>3 minutes). I tried caching it and that also didn't seem to decrease latency. I found that very weird, since including a large file at gemini.google.com and asking something is much faster, usually returning an answer in less than 30s. Why is it that a request containing a large PDF takes too long in the API? Any tips on how to optimize the speed?

#### Environment details

- Programming language: Python
- OS: MacOS
- Language runtime version: 3.11
- Package version: 1.12.1

#### Steps to reproduce

```
chunk_1 = gemini_client.files.get(name="large_doc_part_1.pdf")
chunk_2 = gemini_client.files.get(name="large_doc_part_2.pdf")
response = gemini_client.models.generate_content(
model="gemini-2.5-pro-preview-05-06",
contents=[chunk_1, chunk_2, "Tell me something interesting about this document"],
)
```

Guía de contribución

Abrir la guía de contribución

Evaluación

Este issue todavía no se ha evaluado.

Recibe los nuevos issues en tu correo

Un resumen breve de issues de GitHub para principiantes.