docling-project / docling-project/docling

Support for HunyuanOCR

Open
#3,242 0 comments 0 reactions 0 assignees View on GitHub
enhancement triage/close-stale
Dominant language
Python
Stars
66.4k
Forks
4.8k
Avg merge
2d 21h
Merged PRs (30d)
84

Description

### Requested feature

I've tested this model with [llama.cpp](https://github.com/ggml-org/llama.cpp), it's fast & accurate.

GGUF: https://huggingface.co/ggml-org/HunyuanOCR-GGUF

My tests did include (in Spanish):

- Text and tables: "convert to markdown with tables"
- Diagram with text: "convert to mermaid diagram" or "convert to graphviz's dot"

### Alternatives

- [GLM-OCR](https://huggingface.co/ggml-org/GLM-OCR-GGUF)

---

PD: Docling's default OCR model doesn't support Spanish language text.

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.