r/LocalLLaMA · · 1 min read

OCR, granite-docling-258m vs granite-docling-2stage-258m: has anyone actually noticed any improvements?

Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.

Granite Docling 2stage builds upon the Granite Docling, but introduces a key modifications: it builds a dynamic prompt that precomputes layout objects found within a page, making it more robust on out of distribution data.

What do you think?

submitted by /u/Wise_Stick9613
[link] [comments]

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from r/LocalLLaMA