aws-samples / aws-samples/amazon-textract-textractor
Large PDF response processing is slow
Open
enhancement
latency
- Dominant language
- Jupyter Notebook
- Stars
- 493
- Forks
- 163
- PR merge metrics
- No merged PRs in 30d
Description
When processing large PDFs, processing the response after Textract has generated it can be noticeably slow. We should profile the response parser to identify the bottlenecks.
This seems to be linked to the TABLES feature.
Contributor guide
Assessment
This issue has not been assessed yet.