Support of importing PDFs with multiple papers
Open
Nobody has claimed this yet.
- Dominant language
- JavaScript
- Stars
- 64
- Forks
- 8
- PR merge metrics
- No merged PRs in 30d
Description
A PDF may contain multiple papers. This should be supported by the importer somehow. The user should not have to split the PDF by himself.
Example
Here also the direct link to the PDF:
http://domino.watson.ibm.com/library/CyberDig.nsf/papers/B7ED36FAF73949A4852581E7006B2A55/$File/rc25670.pdf
Contributor guide
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Research direction
Use the linked PDF as a reproduction case and inspect the importer entry points to determine how a document containing multiple papers is currently handled. Define done as importing each paper from the PDF without requiring the user to split it manually, then add coverage for the example case.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- javascript
- Domain
- tooling
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 25/100