OCR, granite-docling-258m vs granite-docling-2stage-258m: has anyone actually noticed any improvements?
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
Granite Docling 2stage builds upon the Granite Docling, but introduces a key modifications: it builds a dynamic prompt that precomputes layout objects found within a page, making it more robust on out of distribution data.
What do you think?
[link] [comments]
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.