docling-project / docling-project/docling

Use HybridChunker from markdown text instead of DoclingDocument

Open
#2,000 2 comments 0 reactions 0 assignees View on GitHub
question triage/close-stale
Dominant language
Python
Stars
66.4k
Forks
4.8k
Avg merge
2d 21h
Merged PRs (30d)
84

Description

### Question

I am converting a PDF document into markdown. Afterwards I need to do some adjustments in the generated markdown text before splitting it into chunks.

Is it somehow possible to use the HybridChunker with the input of a markdown text instead of the docling document?

Thanks in advance for a short feedback!

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.