Dots.Ocr
OCR ProcessingComputer Vision

Dots.Ocr

Multilingual Document Layout Parsing in a Single Vision-Language Model

Open Source

About

dots.ocr Designed for universal accessibility, it possesses the capability to recognize virtually any human script. Beyond achieving state-of-the-art (SOTA) performance in standard multilingual document parsing among models of comparable size, dots.ocr-1.5 excels at converting structured graphics (e.g., charts and diagrams) directly into SVG code, parsing web screens and spotting scene text.

Open Source Health

Not enough history
Stars
9,107
Forks
806
License
MIT
Last commit
7 months ago
Python

Related Categories