docling-project / docling-project/docling

Use openai-like vlm to ocr

Open
#1,894 2 comments 0 reactions 0 assignees View on GitHub
question triage/close-stale
Dominant language
Python
Stars
66.4k
Forks
4.8k
Avg merge
2d 21h
Merged PRs (30d)
84

Description

### Question

I tried to use openai-like vlm to ocr, and I wanted to extract images and contexts in images in PDF and transferred to markdown , however I met with error "Can not append a child with children" and the context and image could not be extracted.
I used PictureDesctiptionVLMoption, PDFpipelineoption and md_export_kwargs to improve the ocr ability,but the outcome was bad . Do I have to change the prompt or others to extract

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.