docling-project / docling-project/docling
How can we extract the content of a PDF page by page using `docling` and convert it into Markdown format?
Open
question
- Dominant language
- Python
- Stars
- 66.4k
- Forks
- 4.8k
- Avg merge
- 2d 21h
- Merged PRs (30d)
- 84
Description
I am working with PDF files and would like to extract their content, one page at a time, and convert it into Markdown format for easier manipulation and display. Specifically, I am looking to use the `docling` tool for this task.
Could someone provide guidance or an example on how to accomplish this, or share any relevant configurations or steps needed to get the content in Markdown format?
Contributor guide
Assessment
This issue has not been assessed yet.