camelot-dev / camelot-dev/camelot
pdf not parsing properly
- Dominant language
- Python
- Stars
- 3.8k
- Forks
- 546
- Avg merge
- 3d 17h
- Merged PRs (30d)
- 3
Description
I wrote some python code to parse PDF files in my system(mac os). In my case, there can be multiple tables on a page in a PDF file. Camelot can parse all tables perfectly in my system. But when I try to run the same code on the server (AWS EC2 Ubuntu) Camelot can only able to parse only a few tables.
Contributor guide
No contributing guide indexed for this repository
Research direction
The report names no source files, tests, PDF sample, dependency versions, or reproducible command. Start by collecting the macOS and Ubuntu EC2 environments and a minimal PDF/code example; done means identifying a reproducible difference and confirming consistent table extraction across both environments.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python
- Domain
- data
- Issue type
- Bug
- Difficulty
- 4/5
- Estimated time
- 3-5 days
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 20/100