camelot-dev / camelot-dev/camelot

pdf not parsing properly

Open
#316 0 comments 0 reactions 0 assignees View on GitHub
bug
Dominant language
Python
Stars
3.8k
Forks
546
Avg merge
3d 17h
Merged PRs (30d)
3

Description

I wrote some python code to parse PDF files in my system(mac os). In my case, there can be multiple tables on a page in a PDF file. Camelot can parse all tables perfectly in my system. But when I try to run the same code on the server (AWS EC2 Ubuntu) Camelot can only able to parse only a few tables.

Contributor guide

No contributing guide indexed for this repository

Research direction

The report names no source files, tests, PDF sample, dependency versions, or reproducible command. Start by collecting the macOS and Ubuntu EC2 environments and a minimal PDF/code example; done means identifying a reproducible difference and confirming consistent table extraction across both environments.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
data
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.