docling-project / docling-project/docling

Picture description with text context

Open
#2,321 1 comment 1 reaction 0 assignees View on GitHub
enhancement triage/close-stale
Dominant language
Python
Stars
66.4k
Forks
4.8k
Avg merge
3d 4h
Merged PRs (30d)
98

Description

### Requested feature
I would like to see an option to provide surrounding text context to image or picture enhancement. I think annotation will be better if VLM has some context in which image is provided. So the model is not getting only the image but for example the section which image is embedded.

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.