aboutcode-org / aboutcode-org/scancode-toolkit
Consider dropping pdfminer and use XPDF to extract text from PDF
オープン
new feature
- 主要言語
- Python
- スター
- 2.6k
- フォーク
- 791
- 平均マージ
- 1日 12時間
- マージ済み PR(30日)
- 5
説明
# Short Description
pdfminer is both slow and has been the source of more than a few issues in the past. Xpdf is C code and os-specific but the pdftotext command may be just enough of what we need:
http://www.xpdfreader.com/download.html
It comes with pre-built command line tools for Linux, Windows and Mac
## Possible Labels
- new feature
## Select Category
- Enhancement [x]
- Add License/Copyright []
- Scan Feature []
- Packaging []
- Documentation []
- Expand Support []
- Other []
コントリビューションガイド
評価
この issue はまだ評価されていません。